If you read the page, Opus is now significantly better than Astra while also being cheaper and having more performance headroom available.
If you read the page, Opus is now significantly better than Astra while also being cheaper and having more performance headroom available.
I read the page. It seems like a marginal improvement.
Let's wait for independent benchmarks at least
the benchmarks provided are already from independent organizations:
Terminal-Bench 4.0 - Stanford & Laude Institute (with funding from all of the AI companies)
FrontierCode v1.1 - Cognition
CursorBench - Cursor (now SolarBoringSpaceXAI I believe)
GDPVal-AA - Artificial Analysis
AutomationBench - Zapier
Humanity's Last Exam - CAIS and Scale AI
Terminal-Bench-Science - Stanford, Laude, Ai2, Allen Institute
OSWOrld - XLANG Lab @ the University of Hong Kong
Chartography - Surge AI
https://artificialanalysis.ai/models/releases/claude-opus-5-...