ANTHROPIC · MODEL RELEASE TRACKER
Claude Opus 5
Claude Opus 5 is an Anthropic model announced as available on 2026-09-16. Anthropic describes it as a thoughtful, proactive model that approaches the frontier intelligence of Claude Fable 5 at roughly half the price, with especially strong results on coding and knowledge-work benchmarks and meaningful gains over Opus 4.8. It is the new default on Claude Max and the strongest model on Claude Pro, and Anthropic notes it remains behind Mythos 5 on cybersecurity tasks.CURRENT SNAPSHOT5/5 DIMENSIONS WITH DATA
The dimensions that change the decision.
Input price$5 USD / 1M input tokens; Claude API standard, global routing
Output price$25 USD / 1M output tokens; Claude API standard, global routing
Cached input price$0.5 USD / 1M cached input tokens; Claude API standard, global routing
Pricing and cost-efficiencyAnthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.
Context window1,000,000 tokens
Input modalitiesText, images
Output modalitiesText
Benchmarks (coding & knowledge work)Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.
SWE-Serve inference-engineering evaluationIn NVIDIA’s SWE-Serve evaluation, Claude Opus 5 is listed among top performers with a reported pass@1 of 75% on the evaluated inference-engineering tasks.
Benchmark performance on coding and knowledge-work evaluationsAnthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).
AvailabilityActive on Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Active (legacy)
RELEASE TIMELINE
Published, source-backed release events only.
VERIFIABLE FACTS
Every value stays attached to a source and date.
RELEASE · Release dateDEVELOPER CLAIM
2026-07-24
CONTEXT WINDOW · Context windowDEVELOPER CLAIM
1,000,000 tokens
CAPABILITIES · Maximum outputDEVELOPER CLAIM
128K tokens
MODALITIES · Input modalitiesDEVELOPER CLAIM
Text, images
MODALITIES · Output modalitiesDEVELOPER CLAIM
Text
AVAILABILITY · AvailabilityDEVELOPER CLAIM
Active on Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Active (legacy)
PRICING · Input priceDEVELOPER CLAIM
$5 USD / 1M input tokens; Claude API standard, global routing
PRICING · Output priceDEVELOPER CLAIM
$25 USD / 1M output tokens; Claude API standard, global routing
PRICING · Cached input priceDEVELOPER CLAIM
$0.5 USD / 1M cached input tokens; Claude API standard, global routing
RELEASE · Release / AvailabilityDEVELOPER CLAIM
Announced available on 2026-09-16 (Anthropic announcement: 'Claude Opus 5 is available today').
OTHER · Performance vs Claude Fable 5DEVELOPER CLAIM
Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.
PRICING · Pricing and cost-efficiencyDEVELOPER CLAIM
Anthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.
BENCHMARKS · Benchmarks (coding & knowledge work)DEVELOPER CLAIM
Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.
LIMITATIONS · Cybersecurity performance (limitation)DEVELOPER CLAIM
Anthropic states Opus 5 remains behind Mythos 5 on cybersecurity tasks.
OTHER · Performance vs Opus 4.8DEVELOPER CLAIM
Reported to provide greatly improved performance for the same cost as its predecessor Opus 4.8.
OTHER · Product placement / Default modelDEVELOPER CLAIM
Set as the new default model on Claude Max and described as the strongest model on Claude Pro.
BENCHMARKS · SWE-Serve inference-engineering evaluationDEVELOPER CLAIM
In NVIDIA’s SWE-Serve evaluation, Claude Opus 5 is listed among top performers with a reported pass@1 of 75% on the evaluated inference-engineering tasks.
BENCHMARKS · Benchmark performance on coding and knowledge-work evaluationsDEVELOPER CLAIM
Anthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).
LIMITATIONS · Relative performance on cybersecurity tasksDEVELOPER CLAIM
Anthropic states Opus 5 remains behind Mythos 5 on cybersecurity tasks.
CAPABILITIES · Recovery of best-supported assignment when constraints are formalizedINDEPENDENTLY SUPPORTED
In a controlled research task, when constraints were stated directly, Claude Opus 5 recovered the best-supported assignment in 96% of cases (reported by the arXiv study).
LIMITATIONS · Sensitivity to scientific prose and biological priorsINDEPENDENTLY SUPPORTED
The arXiv study reports that in scientific prose Opus 5 often prefers the biologically expected assignment; when that biological preference is removed, recovery of the better-supported assignment increases from 27% to 79%, and reaches 92% when the Biology-favored record is accompanied by a formalization request and explicit paired-design cue.
WHAT CHANGED
Stored passport versions, without reconstructed history.
BENCHMARKS · Benchmark performance on coding and knowledge-work evaluations+ Anthropic reports Opus 5 as the new state-of-the-art on coding and knowledge-work evaluations (e.g., surpasses all models on Frontier-Bench v0.1, more than doubles Opus 4.8’s performance at lower cost per task; on CursorBench 3.2 it reaches within 0.5% of Fable 5’s peak at about half the cost per task; on ARC-AGI 3 its score is ~3× the next-best model; on Zapier AutomationBench its pass rate is ~1.5× the next-best for the same cost; on OSWorld 2.0 it outperforms every other model at a given cost).
LIMITATIONS · Relative performance on cybersecurity tasks+ Anthropic states Opus 5 remains behind Mythos 5 on cybersecurity tasks.
PRICING · Pricing and cost-efficiency+ Anthropic positions Opus 5 as delivering greatly improved performance for the same cost as its predecessor Opus 4.8 and says it comes close to the intelligence of Claude Fable 5 at about half the price.
PRICING · Pricing (relative)− Reported to cost about half the price of Claude Fable 5.
CAPABILITIES · Recovery of best-supported assignment when constraints are formalized+ In a controlled research task, when constraints were stated directly, Claude Opus 5 recovered the best-supported assignment in 96% of cases (reported by the arXiv study).
LIMITATIONS · Sensitivity to scientific prose and biological priors+ The arXiv study reports that in scientific prose Opus 5 often prefers the biologically expected assignment; when that biological preference is removed, recovery of the better-supported assignment increases from 27% to 79%, and reaches 92% when the Biology-favored record is accompanied by a formalization request and explicit paired-design cue.
BENCHMARKS · SWE-Serve inference-engineering evaluationListed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.→In NVIDIA’s SWE-Serve evaluation, Claude Opus 5 is listed among top performers with a reported pass@1 of 75% on the evaluated inference-engineering tasks.
AVAILABILITY · Availability+ Active on Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Active (legacy)
RELEASE · Release / AvailabilityMeasurement units or comparison conditions updated
CONTEXT WINDOW · Context window+ 1,000,000 tokens
LIMITATIONS · Cybersecurity performance (limitation)Measurement units or comparison conditions updated
MODALITIES · Input modalities+ Text, images
CAPABILITIES · Maximum output+ 128K tokens
MODALITIES · Output modalities+ Text
OTHER · Performance vs Claude Fable 5Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.→Described as coming close to the frontier intelligence of Claude Fable 5 while costing about half as much.
OTHER · Performance vs Opus 4.8Reported to provide greatly improved performance for the same cost as its predecessor Opus 4.8.→Reported to provide greatly improved performance for the same cost as its predecessor Opus 4.8.
PRICING · Cached input price+ $0.5 USD / 1M cached input tokens; Claude API standard, global routing
PRICING · Input price+ $5 USD / 1M input tokens; Claude API standard, global routing
PRICING · Output price+ $25 USD / 1M output tokens; Claude API standard, global routing
OTHER · Product placement / Default modelSet as the new default model on Claude Max and described as the strongest model on Claude Pro.→Set as the new default model on Claude Max and described as the strongest model on Claude Pro.
PRICING · Pricing (relative)Reported to cost about half the price of Claude Fable 5.→Reported to cost about half the price of Claude Fable 5.
RELEASE · Release date+ 2026-07-24
BENCHMARKS · Benchmarks (coding & knowledge work)Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.→Described as 'the new state-of-the-art' on coding and knowledge-work evaluations such as Frontier-Bench and GDPval-AA.
OTHER · SWE-Serve leaderboardListed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.→Listed among top performers on SWE-Serve; a mini-swe-agent evaluation reported Claude Opus 5 at 75% mean pass@1 in the referenced tasks.