ResearchOctober 05, 2026
27
🌶️
Opus 5.5 Outperforms Human Researchers
TasteVal benchmarks AI research taste: Opus 5.5 beats human experts with a 2.3x compute multiplier.
#benchmark#research taste#Anthropic#Opus#AI progress

🔥 What happened
Opus 5.5 beats human experts on TasteVal, a new benchmark for experimental research taste. The model matches expert performance with 2.3x less compute and costs only 1/30 per run.
💡 Why it matters
Frontier models' compute multipliers now double every 3 months – a huge leap from 14 months before. This means AI progress is accelerating fast, drastically revising forecasts for future models upward.
⚡ Our take
If AI has better research taste than humans, who needs us? Experts should upskill – or focus on interpreting results.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions — when in doubt, read the linked original source.
Deep Dives & Similar Intelligence
Opus 5.5 Under Constant Watch
ResearchSeptember 29, 2026

Zhipu's GLM-5.3 Builds Real Exploits
Ethics & SecuritySeptember 30, 2026

Microsoft and Meta Slam the Brakes on Claude
Business & TrendsOctober 05, 2026

Anthropic's Soul-Searching for Claude
Ethics & SecurityOctober 02, 2026

Sonnet 5.5 Ambushes Opus as OpenAI Teases GPT-6
AI modelsSeptember 29, 2026