@marc.kaz · Marc Kaz
Saved 2026-06-20 · Posted 2026-06-18 · Status: New
Nemotron-3.5-ASR delivers real-time streaming ASR across 40+ languages — and it runs entirely on CPU.
Ships with:
• Only 0.6B parameters
• Real-time streaming output
• 2.5x faster than official Nemo runtime (same accuracy)
• Fully offline capable
• Easy integration into local agent pipelines
Another strong small model for on-device/local AI stacks.
👉 https://huggingface.co/nvidia/nemotron-3.5-asr-streaming-0.6b
Who’s running local speech recognition today? Drop a 🔥
Content ideas (0)
No ideas generated yet. Run /instagram-sync ideate from Claude Code to create some.
Comments (15)
So can someone use this on low level hardware? Like without a GPU?
its really good. recognizes multiple languages, even when switching language mid sentence.
🚨 NVIDIA JUST DROPPED A 0.6B CPU-ONLY SPEECH RECOGNITION MODEL
Nemotron-3.5-ASR delivers real-time streaming ASR across 40+ languages — and it runs entirely on CPU.
Ships with:
• Only 0.6B parameters
• Real-time streaming output
• 2.5x faster than official Nemo runtime (same accuracy)
• Fully offline capable
• Easy integration into local agent pipelines
Another strong small model for on-device/local AI stacks.
👉 https://huggingface.co/nvidia/nemotron-3.5-asr-streaming-0.6b
Who’s running local speech recognition today? Drop a 🔥
You are the best Ai influencer
how is the goat finding these 🔥🙌
Brooo check Mothrag.com new opensource commercial multihop retrival. 😱
nemotron 3.5 vs whisper?
Foh-ee languages
WhatsApp guyz🔥🔥🔥👏 you have hot news everytime
Yo @marc.kaz ! This is actually a good model. I added it to my app LlamaPal. Do check it out. Its on Playstore and it's free.
What is that 241 ms? If thats the actual latency, im not sure its worth it
@sami.p_1707 its working better than whisper for other languages (tested hindi)
How does it compare to Whisper large model?
you guys have tried it, how good it's comparable to grow super whisper?
Why is this any better than existing non NN text to speech tools?