Tools & ProjectsSeptember 26, 2026
-38
🧊
OpenAI's Ultrafast Mode Goes Live
OpenAI preps wider Ultrafast API rollout at 750 tokens/s via Cerebras.
#OpenAI#API#Cerebras#inference

🔥 What happened
OpenAI is gearing up for a broad rollout of its "Ultrafast" API mode around DevDay on September 29. A hidden speed selector — Standard, Fast, Ultrafast — has already surfaced in the API docs.
💡 Why it matters
OpenAI previewed Ultrafast at up to 750 output tokens per second, roughly 14× faster than Standard, powered by Cerebras. Developers could soon pick latency per workload, reserving pricey inference for jobs where speed actually drives revenue.
⚡ Our take
This isn't a feature, it's a pricing decision. Teams that default to Ultrafast will burn budget on workloads where 200 ms never mattered.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions — when in doubt, read the linked original source.
Deep Dives & Similar Intelligence

OpenAI's $500 ChatGPT Pro Max Leaks
Business & TrendsSeptember 24, 2026
ki-daily.
OpenAI Cracks 100 Math Problems
ResearchSeptember 22, 2026

OpenAI Pulls Model Over Deception Fears
Ethics & SecuritySeptember 28, 2026

AI Labs Are Losing Control of Their Agents
Ethics & SecuritySeptember 27, 2026

OpenAI Knew About Book Piracy
Ethics & SecuritySeptember 27, 2026