Trendora

Speech-to-Speech

Assess

Techniques

Technology that transforms spoken input directly into spoken output.

Why it's here

Placed in Assess: 3 article(s) of evidence from 1 source(s), led by product launches, with 1 in the last 30 days. Confidence 42%.

Evidence (3)

  • 7Hugging Face Blog·7/15/2026research
    Hume AI unveils Real World VoiceEQ for evaluating voice AI quality

    Hume AI introduced Real World VoiceEQ, a benchmark for measuring the human quality of voice AI across speech recognition, text-to-speech, speech-to-speech, and speech understanding. Built from over 1 million human ratings, it evaluates more than 40 models across 15+ dimensions and suggests current voice systems are increasingly specialized rather than having a single best performer.

  • 6Hugging Face Blog·5/27/2026product_launch
    Reachy Mini can now run fully locally

    Hugging Face announced that Reachy Mini’s conversation stack can now run entirely on local hardware instead of sending audio to a server. The setup uses a cascaded speech-to-speech pipeline with llama.cpp, Silero VAD, Parakeet-TDT 0.6B v3 for STT, and Qwen3-TTS, enabling private offline conversations with the robot.

  • 7Hugging Face Blog·3/24/2026framework_update
    Hugging Face Blog Introduces EVA for Voice Agent Evaluation

    ServiceNow AI researchers present EVA, an end-to-end framework for evaluating conversational voice agents across both task accuracy and spoken interaction quality. The framework outputs two scores, EVA-A and EVA-X, and ships with an airline dataset plus benchmark results for cascade and audio-native systems.