ResearchSeptember 23, 2026
-36
🧊

Kyutai Teaches Speech Models to Do Math Out Loud

Kyutai Voice of Reason solves spoken math via RL: 27% to 77%.

#Kyutai#Speech#Reinforcement Learning#Open Weights#Math
Kyutai bringt Sprachmodellen das Rechnen bei
Share Article
šŸ”„ What happened Kyutai released Voice of Reason, two open-weight speech-to-speech models that solve math problems out loud. No transcription step, no separate text LLM. On spoken GSM8K, accuracy jumps from 27.3% to 77.1%. šŸ’” Why it matters The recipe: supervised fine-tuning plus reinforcement learning on 16 H100 GPUs. One detail is critical – without temperature correction, accuracy collapses to 12.3%. Both 9B checkpoints run on a single H100, making self-hosting realistic. ⚔ Our take Speech-native reasoning without a text detour is a genuine paradigm shift. Anyone still betting on cascaded pipelines is sacrificing latency, paralinguistic cues, and soon, relevance.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions — when in doubt, read the linked original source.