Tech Meridian← ALL MODELS
PROMY MERIDIAN

XAI · MODEL RELEASE TRACKER

Grok 4.7

Grok 4.7 is a Grok family model released by xAI and made available on Amazon Bedrock. xAI positions it as their most capable model for coding and knowledge work, emphasizing endurance on long tasks, stronger self-verification, and support for long context windows and configurable reasoning effort.

CURRENT SNAPSHOT3/5 DIMENSIONS WITH DATA

The dimensions that change the decision.

PRICING

Not established from the available sources.

CONTEXT WINDOW
Context windowGrok 4.7 offers a 500K token context window.
MODALITIES

Not established from the available sources.

BENCHMARKS
CursorBench 4.0 price-performancexAI reports Grok 4.7 is at the frontier in price-performance on CursorBench 4.0, a benchmark that stresses longer-running coding tasks.
GDPval and AA Briefcase performancexAI reports Grok 4.7 improves on Grok 4.6 on GDPval and AA Briefcase (benchmarks for professional tasks) and performs comparably to other frontier models.
Artificial Analysis Intelligence Index (v4.3.2) scoreAn independent report gives Grok 4.7 a combined score of 46 on the Artificial Analysis Intelligence Index (v4.3.2), placing it mid-pack versus competitors (Claude Fable 5.1 and GPT-6 scored 53).
Terminal-Bench 4.0 (agentic coding)Reported Terminal-Bench 4.0 agentic-coding score of 26% (lower than certain competitor models reported in the same article).
Reported benchmark gainsxAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.
AVAILABILITY
Availability and APIsGrok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.

RELEASE TIMELINE

Published, source-backed release events only.

2 SOURCES · IMPORTANCE 8.0

SpaceXAI releases Grok 4.7 — faster, cheaper model for coding and knowledge work

SpaceXAI released Grok 4.7, a new base-model upgrade positioned as its most capable model for coding and knowledge work; it reportedly uses a larger base model, longer RL training focused on long-running tasks, and an entirely new safeguard stack. The company says Grok 4.7 matches the price and speed of Grok 4.6 while improving verification, longer-context handling, document/presentation generation, benchmark performance (CursorBench 4.0, GDPval, AA Briefcase), and safety metrics (LatchBio biosafety 62.4%, HackerBench v0.3 allowing 3.3% risky prompts); it is available in Cursor, Grok Build, the Grok API, third-party harnesses and cloud platforms, priced from $2 per million input tokens and $6 per million output tokens, with an optional faster (2x) variant at double the price.

→

VERIFIABLE FACTS

Every value stays attached to a source and date.

BENCHMARKS · CursorBench 4.0 price-performanceDEVELOPER CLAIM

xAI reports Grok 4.7 is at the frontier in price-performance on CursorBench 4.0, a benchmark that stresses longer-running coding tasks.

BENCHMARKS · GDPval and AA Briefcase performanceDEVELOPER CLAIM

xAI reports Grok 4.7 improves on Grok 4.6 on GDPval and AA Briefcase (benchmarks for professional tasks) and performs comparably to other frontier models.

AVAILABILITY · Availability and APIsDEVELOPER CLAIM

Grok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.

RELEASE · Training and base modelDEVELOPER CLAIM

xAI reports Grok 4.7 uses a new, larger base model and was trained with a longer reinforcement-learning run over a harder mix of tasks, deliberately weighted toward problems that take many hours to complete.

CAPABILITIES · Grok Bot harness and conversational gainsDEVELOPER CLAIM

xAI reports Grok 4.7 was trained to natively understand the Grok Bot harness, which the company credits for improvements in conversational tasks and general knowledge work.

BENCHMARKS · Reported benchmark gainsDEVELOPER CLAIM

xAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.

WHAT CHANGED

Stored passport versions, without reconstructed history.

Passport updated15 facts
AVAILABILITY · Availability and APIs+ Grok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.
CAPABILITIES · Configurable reasoning effort+ The model supports configurable reasoning effort at four levels: low, medium, high, and xhigh.
CONTEXT WINDOW · Context window+ Grok 4.7 offers a 500K token context window.
CAPABILITIES · Grok Bot harness and conversational gains+ xAI reports Grok 4.7 was trained to natively understand the Grok Bot harness, which the company credits for improvements in conversational tasks and general knowledge work.
CONTEXT WINDOW · context handling− Described by xAI as better at managing longer context than prior versions (no numeric context-window size provided).
CAPABILITIES · Primary use cases+ xAI positions Grok 4.7 as built for coding, long-running agents, and professional knowledge work (e.g., legal, medical, financial workflows).
BENCHMARKS · Reported benchmark gains+ xAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.
SAFETY · Self-verification behavior+ The model checks and verifies its own output more carefully before moving on; xAI states this self-verification helps the model fail less catastrophically on long trajectories.
RELEASE · Training and base model+ xAI reports Grok 4.7 uses a new, larger base model and was trained with a longer reinforcement-learning run over a harder mix of tasks, deliberately weighted toward problems that take many hours to complete.
Passport created8 facts