Tech Meridian ← ENTITY INDEX
RU

MODEL · ENTITY #3707

Claude Opus 4.8

Related event timeline, sources and context from the news index.

EVENT TIMELINE

2

MODELS · 4 SOURCES · Anthropic · TechCrunch AI · The Verge AI · The Decoder

Anthropic releases Claude Opus 5 — lower-cost model claiming near‑frontier performance

Anthropic announced Claude Opus 5, a new model positioned as a cost‑efficient successor to Opus 4.8 and the new default on Claude Max (and the strongest on Claude Pro). Anthropic says Opus 5 matches or exceeds prior models on many coding, knowledge‑work and scientific benchmarks (Frontier‑Bench, GDPval‑AA, CursorBench, ARC‑AGI, Zapier AutomationBench, OSWorld) at lower cost per task while remaining behind Mythos 5 on security and biology frontier tasks; the company also reports improved alignment and safety in pre‑deployment audits and links a System Card for more details.

8.0

RESEARCH · 1 SOURCE · arXiv cs.AI

arXiv paper introduces SAFE benchmark to test whether frontier models seek safety evidence before acting

The paper (arXiv:2609.17865v1) introduces SAFE, a controlled benchmark where models decide whether to retrieve optional safety-relevant evidence that varies in cost, probability, severity, and presentation. Evaluating GPT-5.5, o3, Claude Opus 4.8, and Claude Sonnet 4.6, the authors find distinct evidence-acquisition policies (Opus inspects by default, o3 skips most, GPT-5.5 and Sonnet intermediate), strong sensitivity to severity and retrieval cost, weaker sensitivity to probability, and a mismatch between stated rationales and actual causal influences.

7.0