Sonnet 5.5 Hits 70.6% on Terminal-Bench 4.0 for $2/$10 per Million
Claude Sonnet 5.5 scores 70.6% on Terminal-Bench 4.0 vs Sonnet 5's 10.3%, at the same $2/$10 per-million pricing but up to 30% cheaper per task.
5 posts
Claude Sonnet 5.5 scores 70.6% on Terminal-Bench 4.0 vs Sonnet 5's 10.3%, at the same $2/$10 per-million pricing but up to 30% cheaper per task.
Opus 5.5 ships at $4/$20 per MT with $0.20 cache reads, beating GPT-6 Astra at 20% of the cost per task.
Anthropic shipped Opus 4.7 on April 16, 2026, with a seven-point SWE-bench jump, the 1M context window now generally available with no premium, and a new task budget primitive for agent loops.
Anthropic released Opus 4.6 on February 5, 2026, with a 1M token context beta, agent teams, adaptive thinking, and developer effort controls — all at the same price as 4.5.
Anthropic's latest model achieves state-of-the-art results in agentic coding and brings meaningful improvements across reasoning, mathematics, and everyday tasks.