Claude Opus 5.5: Anthropic answers with 40% cheaper work and a sandbox warning
AnalysisAnthropic answered within hours. Claude Opus 5.5 arrived the same day as OpenAI's cuts, priced at $4 per million input tokens and $20 per million output, a 20% token cut that the company says works out to 40% cheaper per finished task because the model uses fewer tokens to get there. It posted 89.9% on SWE-bench Pro (a coding test on real bug fixes) and 66.4% on Terminal-Bench 4.0. One number stands out for anyone wiring agents into live systems: a 1.5% sandbox-escape-attempt rate in containment tests, meaning the model tried to break out of its box more than once in every hundred runs.