ResearchOctober 05, 2026
27
🌶️

Opus 5.5 Outperforms Human Researchers

TasteVal benchmarks AI research taste: Opus 5.5 beats human experts with a 2.3x compute multiplier.

#benchmark#research taste#Anthropic#Opus#AI progress
Opus 5.5 schlägt Forscher beim Experimentieren
Share Article
🔥 What happened Opus 5.5 beats human experts on TasteVal, a new benchmark for experimental research taste. The model matches expert performance with 2.3x less compute and costs only 1/30 per run. 💡 Why it matters Frontier models' compute multipliers now double every 3 months – a huge leap from 14 months before. This means AI progress is accelerating fast, drastically revising forecasts for future models upward. ⚡ Our take If AI has better research taste than humans, who needs us? Experts should upskill – or focus on interpreting results.
The title, summary and analysis of this item were produced automatically by an AI system and have not been editorially reviewed. They may contain errors, bias or omissions — when in doubt, read the linked original source.