Tech Meridian← ALL MODELS
PROMY MERIDIAN

ANTHROPIC · MODEL RELEASE TRACKER

Claude Opus 5

Claude Opus 5 is an Anthropic model announced as available on 2026-09-16. Anthropic describes it as a thoughtful, proactive model that approaches the frontier intelligence of Claude Fable 5 at roughly half the price, with especially strong results on coding and knowledge-work benchmarks and meaningful gains over Opus 4.8. It is the new default on Claude Max and the strongest model on Claude Pro, and Anthropic notes it remains behind Mythos 5 on cybersecurity tasks.

CURRENT SNAPSHOT5/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING
Input price$5 USD / 1M input tokens; Claude API standard, global routing
Output price$25 USD / 1M output tokens; Claude API standard, global routing
Cached input price$0.5 USD / 1M cached input tokens; Claude API standard, global routing
Pricing and cost-efficiencyAnthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.
CONTEXT WINDOW
Context window1,000,000 tokens
MODALITIES
Input modalitiesText, images
Output modalitiesText
BENCHMARKS
Benchmarks (coding & knowledge work)Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.
SWE-Serve inference-engineering evaluationIn NVIDIA’s SWE-Serve evaluation, Claude Opus 5 is listed among top performers with a reported pass@1 of 75% on the evaluated inference-engineering tasks.
Benchmark performance on coding and knowledge-work evaluationsAnthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).
AVAILABILITY
AvailabilityActive on Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Active (legacy)

RELEASE TIMELINE

Published, source-backed release events only.

6 SOURCES · IMPORTANCE 8.0

Anthropic releases Claude Opus 5 — lower-cost model claiming near‑frontier performance

Anthropic announced Claude Opus 5, a new model positioned as a cost‑efficient successor to Opus 4.8 and the new default on Claude Max (and the strongest on Claude Pro). Anthropic says Opus 5 matches or exceeds prior models on many coding, knowledge‑work and scientific benchmarks (Frontier‑Bench, GDPval‑AA, CursorBench, ARC‑AGI, Zapier AutomationBench, OSWorld) at lower cost per task while remaining behind Mythos 5 on security and biology frontier tasks; the company also reports improved alignment and safety in pre‑deployment audits and links a System Card for more details.

→

VERIFIABLE FACTS

Every value stays attached to a source and date.

OTHER · Performance vs Claude Fable 5DEVELOPER CLAIM

Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.

PRICING · Pricing and cost-efficiencyDEVELOPER CLAIM

Anthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.

BENCHMARKS · Benchmarks (coding & knowledge work)DEVELOPER CLAIM

Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.

BENCHMARKS · Benchmark performance on coding and knowledge-work evaluationsDEVELOPER CLAIM

Anthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).

LIMITATIONS · Sensitivity to scientific prose and biological priorsINDEPENDENTLY SUPPORTED

The arXiv study reports that in scientific prose Opus 5 often prefers the biologically expected assignment; when that biological preference is removed, recovery of the better-supported assignment increases from 27% to 79%, and reaches 92% when the Biology-favored record is accompanied by a formalization request and explicit paired-design cue.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport updated21 facts
BENCHMARKS · Benchmark performance on coding and knowledge-work evaluations+ Anthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).
LIMITATIONS · Relative performance on cybersecurity tasks+ Anthropic states Opus 5 remains behind Mythos 5 on cybersecurity tasks.
PRICING · Pricing and cost-efficiency+ Anthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.
PRICING · Pricing (relative)− Reported to cost about half the price of Claude Fable 5.
CAPABILITIES · Recovery of best-supported assignment when constraints are formalized+ In a controlled research task, when constraints were stated directly, Claude Opus 5 recovered the best-supported assignment in 96% of cases (reported by the arXiv study).
LIMITATIONS · Sensitivity to scientific prose and biological priors+ The arXiv study reports that in scientific prose Opus 5 often prefers the biologically expected assignment; when that biological preference is removed, recovery of the better-supported assignment increases from 27% to 79%, and reaches 92% when the Biology-favored record is accompanied by a formalization request and explicit paired-design cue.
BENCHMARKS · SWE-Serve inference-engineering evaluationListed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.→In NVIDIA’s SWE-Serve evaluation, Claude Opus 5 is listed among top performers with a reported pass@1 of 75% on the evaluated inference-engineering tasks.
Passport updated17 facts
AVAILABILITY · Availability+ Active on Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Active (legacy)
RELEASE · Release / AvailabilityMeasurement units or comparison conditions updated
CONTEXT WINDOW · Context window+ 1,000,000 tokens
LIMITATIONS · Cybersecurity performance (limitation)Measurement units or comparison conditions updated
MODALITIES · Input modalities+ Text, images
CAPABILITIES · Maximum output+ 128K tokens
MODALITIES · Output modalities+ Text
OTHER · Performance vs Claude Fable 5Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.→Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.
OTHER · Performance vs Opus 4.8Reported to provide greatly improved performance for the same cost as its predecessor Opus 4.8.→Reported to provide greatly improved performance for the same cost as its predecessor Opus 4.8.
PRICING · Cached input price+ $0.5 USD / 1M cached input tokens; Claude API standard, global routing
PRICING · Input price+ $5 USD / 1M input tokens; Claude API standard, global routing
PRICING · Output price+ $25 USD / 1M output tokens; Claude API standard, global routing
OTHER · Product placement / Default modelSet as the new default model on Claude Max and described as the strongest model on Claude Pro.→Set as the new default model on Claude Max and described as the strongest model on Claude Pro.
PRICING · Pricing (relative)Reported to cost about half the price of Claude Fable 5.→Reported to cost about half the price of Claude Fable 5.
RELEASE · Release date+ 2026-07-24
BENCHMARKS · Benchmarks (coding & knowledge work)Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.→Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.
OTHER · SWE-Serve leaderboardListed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.→Listed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.
Passport created8 facts