Not established from the available sources.
XAI · MODEL RELEASE TRACKER
Grok 4.7
Grok 4.7 is a Grok family model released by xAI and made available on Amazon Bedrock. xAI positions it as their most capable model for coding and knowledge work, emphasizing endurance on long tasks, stronger self-verification, and support for long context windows and configurable reasoning effort.CURRENT SNAPSHOT3/5 DIMENSIONS WITH DATA
The dimensions that change the decision.
Context windowGrok 4.7 offers a 500K token context window.
Not established from the available sources.
CursorBench 4.0 price-performancexAI reports Grok 4.7 is at the frontier in price-performance on CursorBench 4.0, a benchmark that stresses longer-running coding tasks.
GDPval and AA Briefcase performancexAI reports Grok 4.7 improves on Grok 4.6 on GDPval and AA Briefcase (benchmarks for professional tasks) and performs comparably to other frontier models.
Artificial Analysis Intelligence Index (v4.3.2) scoreAn independent report gives Grok 4.7 a combined score of 46 on the Artificial Analysis Intelligence Index (v4.3.2), placing it mid-pack versus competitors (Claude Fable 5.1 and GPT-6 scored 53).
Terminal-Bench 4.0 (agentic coding)Reported Terminal-Bench 4.0 agentic-coding score of 26% (lower than certain competitor models reported in the same article).
Reported benchmark gainsxAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.
Availability and APIsGrok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.
RELEASE TIMELINE
Published, source-backed release events only.
VERIFIABLE FACTS
Every value stays attached to a source and date.
RELEASE · release / availabilityDEVELOPER CLAIM
Announced and made available by xAI on 2026-09-21; described as "available today" in the announcement.
CAPABILITIES · primary use / specializationDEVELOPER CLAIM
Positioned as xAI's most capable model for coding and knowledge work.
CAPABILITIES · base model and trainingDEVELOPER CLAIM
Built on a new, larger base model compared to Grok 4.6 and trained with a longer reinforcement‑learning run on a harder mix of tasks weighted toward problems that take many hours to complete.
CONTEXT WINDOW · Context windowDEVELOPER CLAIM
Grok 4.7 offers a 500K token context window.
BENCHMARKS · CursorBench 4.0 price-performanceDEVELOPER CLAIM
xAI reports Grok 4.7 is at the frontier in price-performance on CursorBench 4.0, a benchmark that stresses longer-running coding tasks.
BENCHMARKS · GDPval and AA Briefcase performanceDEVELOPER CLAIM
xAI reports Grok 4.7 improves on Grok 4.6 on GDPval and AA Briefcase (benchmarks for professional tasks) and performs comparably to other frontier models.
BENCHMARKS · Artificial Analysis Intelligence Index (v4.3.2) scoreDEVELOPER CLAIM
An independent report gives Grok 4.7 a combined score of 46 on the Artificial Analysis Intelligence Index (v4.3.2), placing it mid-pack versus competitors (Claude Fable 5.1 and GPT-6 scored 53).
BENCHMARKS · Terminal-Bench 4.0 (agentic coding)DEVELOPER CLAIM
Reported Terminal-Bench 4.0 agentic-coding score of 26% (lower than certain competitor models reported in the same article).
AVAILABILITY · Availability and APIsDEVELOPER CLAIM
Grok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.
CAPABILITIES · Primary use casesDEVELOPER CLAIM
xAI positions Grok 4.7 as built for coding, long-running agents, and professional knowledge work (e.g., legal, medical, financial workflows).
CAPABILITIES · Configurable reasoning effortDEVELOPER CLAIM
The model supports configurable reasoning effort at four levels: low, medium, high, and xhigh.
RELEASE · Training and base modelDEVELOPER CLAIM
xAI reports Grok 4.7 uses a new, larger base model and was trained with a longer reinforcement-learning run over a harder mix of tasks, deliberately weighted toward problems that take many hours to complete.
SAFETY · Self-verification behaviorDEVELOPER CLAIM
The model checks and verifies its own output more carefully before moving on; xAI states this self-verification helps the model fail less catastrophically on long trajectories.
CAPABILITIES · Grok Bot harness and conversational gainsDEVELOPER CLAIM
xAI reports Grok 4.7 was trained to natively understand the Grok Bot harness, which the company credits for improvements in conversational tasks and general knowledge work.
BENCHMARKS · Reported benchmark gainsDEVELOPER CLAIM
xAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.
WHAT CHANGED
Stored passport versions, without reconstructed history.
AVAILABILITY · Availability and APIs+ Grok 4.7 is available on Amazon Bedrock, served on the bedrock-runtime endpoint through cross-Region inference profiles, and supports the Responses, Chat Completions, and Converse APIs.
CAPABILITIES · Configurable reasoning effort+ The model supports configurable reasoning effort at four levels: low, medium, high, and xhigh.
CONTEXT WINDOW · Context window+ Grok 4.7 offers a 500K token context window.
CAPABILITIES · Grok Bot harness and conversational gains+ xAI reports Grok 4.7 was trained to natively understand the Grok Bot harness, which the company credits for improvements in conversational tasks and general knowledge work.
CONTEXT WINDOW · context handling− Described by xAI as better at managing longer context than prior versions (no numeric context-window size provided).
CAPABILITIES · Primary use cases+ xAI positions Grok 4.7 as built for coding, long-running agents, and professional knowledge work (e.g., legal, medical, financial workflows).
BENCHMARKS · Reported benchmark gains+ xAI reports gains across published evaluations, including software engineering benchmarks (CursorBench, DeepSWE), multi-hour terminal and office work (Terminal-Bench, AA Briefcase), and domain benchmarks such as EEBench for electrical engineering and evaluations for legal work.
SAFETY · Self-verification behavior+ The model checks and verifies its own output more carefully before moving on; xAI states this self-verification helps the model fail less catastrophically on long trajectories.
RELEASE · Training and base model+ xAI reports Grok 4.7 uses a new, larger base model and was trained with a longer reinforcement-learning run over a harder mix of tasks, deliberately weighted toward problems that take many hours to complete.