Skip to content

Pull requests: NVIDIA-NeMo/Speech

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

Improve MT-Parakeet streaming inference and timestamps ASR Speaker Tasks
#16070 opened Aug 17, 2026 by ipmedenn Collaborator Loading…
6 of 8 tasks
fix(asr): preserve Unicode in transcription JSON ASR community-request waiting-on-customer Waiting on the original author to respond
#16066 opened Aug 15, 2026 by sylvesterkaczmarek Loading…
1 of 3 tasks
Correctly Propagate HF OFFLINE Environment Vars to Container CI core Changes to NeMo Core TTS
#16065 opened Aug 13, 2026 by blisc Collaborator Loading…
1 of 3 tasks
Preserve valid codec subframe before audio EOS TTS
#16059 opened Aug 13, 2026 by vklimkov-nvidia Member Loading…
Fix variable name for audio file path in '_write_to_tar' function community-request waiting-on-maintainers Waiting on maintainers to respond
#16058 opened Aug 13, 2026 by advait-bm Loading…
5 of 8 tasks
Add native THD packed ASR encoders ASR common
#16053 opened Aug 11, 2026 by pzelasko Collaborator 3/3 Loading…
Support chunk microbatching and idxpack for multimodal SpeechLM data common
#16048 opened Aug 10, 2026 by KunalDhawan Collaborator 2/3 Loading…
adding support for multilinugal cache aware model ASR
#16047 opened Aug 10, 2026 by naymaraq Collaborator Loading…
2 of 8 tasks
[Streaming SpeechLM] Add audomodel support ASR
#16045 opened Aug 7, 2026 by stevehuang52 Collaborator Draft
add simulstream backend + qwen reasoning models for ST ASR
#16043 opened Aug 7, 2026 by lilithgrigoryan Collaborator Loading…
8 tasks
ProTip! Updated in the last three days: updated:>2026-08-15.