ResearchSeptember 23, 2026
-36
š§
Kyutai Teaches Speech Models to Do Math Out Loud
Kyutai Voice of Reason solves spoken math via RL: 27% to 77%.
#Kyutai#Speech#Reinforcement Learning#Open Weights#Math

š„ What happened
Kyutai released Voice of Reason, two open-weight speech-to-speech models that solve math problems out loud. No transcription step, no separate text LLM. On spoken GSM8K, accuracy jumps from 27.3% to 77.1%.
š” Why it matters
The recipe: supervised fine-tuning plus reinforcement learning on 16 H100 GPUs. One detail is critical ā without temperature correction, accuracy collapses to 12.3%. Both 9B checkpoints run on a single H100, making self-hosting realistic.
ā” Our take
Speech-native reasoning without a text detour is a genuine paradigm shift. Anyone still betting on cascaded pipelines is sacrificing latency, paralinguistic cues, and soon, relevance.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions ā when in doubt, read the linked original source.
Deep Dives & Similar Intelligence

Xiaomi Crashes the Open-Weights Party
AI modelsSeptember 22, 2026

Qwen Spins Up 3,300 RL Worlds for Cents
ResearchSeptember 24, 2026

NVIDIA Doubles Down on Speaker Diarization
AI modelsSeptember 23, 2026

Google's Agents Learn by Dreaming
ResearchSeptember 19, 2026

Gander Kills the Turn-Taking Era
ResearchSeptember 20, 2026