Stora hopp i kapacitet och kostnad för nya Anthropic modellen. När blev den här modellen färdig för dom internt och började jobba med dom med deras nästa modell? Jag vet inte men skulle inte förvåna om dom suttit på den här sedan april eller ännu tidigare.
Citat:
Claude Opus 5
Anthropic·Jul 24, 2026·2 reasoning variantsClaude Opus 5 sets a new high score on ARC-AGI-3. As of July 24, 2026, Claude Opus 5 (High) is the highest-performing model on ARC-AGI-3, scoring 30.2%. It completed five additional Public Demo environments that no model had previously beaten, demonstrating strong logical reasoning. At Max reasoning effort, Opus 5 scores 97.5% on ARC-AGI-1 and 90.4% on ARC-AGI-2 Semi-Private. This is competitive with previous frontier leaders, though at slightly higher cost. Due to the short testing window, ARC-AGI-3 was evaluated only at High reasoning effort
https://arcprize.org/results/anthropic-claude-opus-5
Citat:
Performance and cost-effectiveness
Claude Opus 5 provides greatly improved performance for the same cost as its predecessor, Opus 4.8. The charts in this section show how performance changes according to the model’s effort setting, which customers can use to optimize for intelligence or conserve tokens for faster and cheaper results.
Opus 5 excels on valuable software engineering tasks. For example, on Frontier-Bench v0.1, Opus 5 surpasses all other models, and more than doubles Opus 4.8’s performance at a lower cost per task. On CursorBench 3.2, at max effort, the model performs within 0.5% of Fable 5’s peak score, but at half the cost per task; it also achieves greater performance at a given cost than all other models on high, xhigh, and max effort.
https://www.anthropic.com/news/claude-opus-5