Trendora

Parakeet-TDT 0.6B v3

Assess

Languages & Frameworks

An automatic speech recognition model used to convert speech to text.

Why it's here

Placed in Assess: 2 article(s) of evidence from 1 source(s), led by product launches, with 0 in the last 30 days. Confidence 33%.

Evidence (2)

  • 6Hugging Face Blog·5/27/2026product_launch
    Reachy Mini can now run fully locally

    Hugging Face announced that Reachy Mini’s conversation stack can now run entirely on local hardware instead of sending audio to a server. The setup uses a cascaded speech-to-speech pipeline with llama.cpp, Silero VAD, Parakeet-TDT 0.6B v3 for STT, and Qwen3-TTS, enabling private offline conversations with the robot.

  • 8Hugging Face Blog·4/28/2026model_release
    NVIDIA Launches Nemotron 3 Nano Omni for Long-Context Multimodal AI

    NVIDIA introduced Nemotron 3 Nano Omni, an omni-modal model for document analysis, speech recognition, long audio-video understanding, and agentic computer use. The model claims strong benchmark results on document, video, audio, and voice tasks, along with higher throughput and efficiency than comparable open multimodal models.