Parakeet-TDT 0.6B v3
AssessLanguages & Frameworks
An automatic speech recognition model used to convert speech to text.
Why it's here
Placed in Assess: 2 article(s) of evidence from 1 source(s), led by product launches, with 0 in the last 30 days. Confidence 33%.
Evidence (2)
- 6Hugging Face Blog·5/27/2026product_launchReachy Mini can now run fully locally
Hugging Face announced that Reachy Mini’s conversation stack can now run entirely on local hardware instead of sending audio to a server. The setup uses a cascaded speech-to-speech pipeline with llama.cpp, Silero VAD, Parakeet-TDT 0.6B v3 for STT, and Qwen3-TTS, enabling private offline conversations with the robot.
- 8Hugging Face Blog·4/28/2026model_releaseNVIDIA Launches Nemotron 3 Nano Omni for Long-Context Multimodal AI
NVIDIA introduced Nemotron 3 Nano Omni, an omni-modal model for document analysis, speech recognition, long audio-video understanding, and agentic computer use. The model claims strong benchmark results on document, video, audio, and voice tasks, along with higher throughput and efficiency than comparable open multimodal models.