NVIDIA Releases Nemotron 3.5 ASR: A 600M-Parameter Cache-Aware Streaming Model Transcribing 40 Language-Locales in Real Time
NVIDIA’s Nemotron Speech workforce has launched Nemotron 3.5 ASR. It is a 600M-parameter streaming Automatic Speech Recognition (ASR) mannequin. A single checkpoint transcribes 40 language-locales in actual time. Punctuation and capitalization are constructed in natively. The mannequin ships as open weights on Hugging Face. The license is OpenMDW-1.1. The structure is a Cache-Aware QuickConformer-RNNT. What…
