Quration Play

tech

In November 2025, Anthropic released a new flagship model that became the first AI system to score above 80% on the SWE-bench Verified software engineering benchmark.

What model did Anthropic release in November 2025 that broke the 80% SWE-bench mark?

Claude Opus 4.5
Claude Opus 4.5 launched on November 24, 2025, scoring 80.9% on SWE-bench Verified and reportedly cutting token usage roughly in half compared to prior models on coding tasks. It rounded out Anthropic's 4.5 model family, following Sonnet 4.5 and Haiku 4.5 earlier that year.

Claude Opus 4.5 is correct — Anthropic released it on November 24, 2025, and it became the first AI model to score above 80% on SWE-bench Verified, a benchmark that tests how well an AI can fix real-world software bugs, reaching 80.9% while reportedly using roughly half the tokens of prior models on comparable coding tasks.

Claude Sonnet 5 and Claude Haiku 5 are both real, capable models, but neither is the specific model that hit this benchmark milestone in November 2025 — Opus, as the largest and most capable model in the family, was the one that pushed past the 80% mark.

Opus 4.5 rounded out Anthropic's '4.5' family of models, following Sonnet 4.5 and Haiku 4.5 earlier in the same year, each aimed at different balances of capability, speed, and cost.

▶ Quration Play

tech

How does GPS actually calculate a device's position?Which region is it?What did Google call this open model family released in March 2025?About how much could this brain-inspired chip cut AI systems' energy use by?What is the most important reason for taking this approach?What is this numeric identifier called?What is its structural cause?What is this Microsoft chip called?What is this model called?Which was invented first — the fax machine or the telephone?Who is widely regarded as the world's first computer programmer?What does the "http" at the start of a web address stand for?

Quration — Quration Play