AI & Tools
Claude Opus 5: Benchmarks, Price and Cost Paradox
Claude Opus 5 benchmarked: first place in the Intelligence Index with 61 points, half the Fable 5 price and the cost paradox of verbosity.
13 min read
Mijo Jurisic
Read more3 articles
Claude Opus 5 benchmarked: first place in the Intelligence Index with 61 points, half the Fable 5 price and the cost paradox of verbosity.
OpenAI's GPT-5.6 line (Luna, Terra, Sol) in a benchmark check and the METR cheating findings: why strong numbers and reward hacking show up together.
Kimi K3 by Moonshot AI benchmarked: the largest open model, ranked 2nd to 4th against Fable 5 and GPT-5.6 Sol, best price-performance ratio.