xAI has launched Grok 4.7, its latest frontier model focused on coding and knowledge work, with improved performance on long running tasks and stronger safeguards.
The model uses a larger base model than Grok 4.6 and underwent a longer reinforcement learning run focused on more difficult tasks that can take hours to complete. xAI said the training improved Grok 4.7’s ability to verify its own work, manage longer context and operate within the Grok Bot environment.
Grok 4.7 improves on its predecessor across the benchmarks published by xAI. It scored 46.3% on CursorBench 4.0 compared with 40.4% for Grok 4.6, while its Terminal Bench 4.0 score rose to 38% from 20.3%.
The model also scored 71% on DeepSWE 1.1 at high reasoning effort, 1,657 on AA Briefcase and 64% on EEBench. xAI’s results put Grok 4.7 ahead of GPT 5.6 Sol on several of those evaluations, though GPT 5.6 Sol and Fable 5.1 remained ahead on others.
The release follows Grok 4.6, which xAI launched in August with a focus on long running agents and complex coding and knowledge work. Grok 4.6 was priced at $2 per million input tokens and $6 per million output tokens.
xAI is keeping the same starting price for Grok 4.7 at $2 per million input tokens and $6 per million output tokens. A faster version offering twice the output speed is available at twice the price.
The company also introduced a new safeguard stack for Grok 4.7. xAI said the model allowed 3.3% of risky dual use prompts through on HackerBench 0.3 and scored 62.4% on a biosafety benchmark from LatchBio.
The company is also providing selected cybersecurity partners with access to the model’s red team capabilities for defensive research.
Grok 4.7 is available through Cursor, Grok Build and the Grok API, as well as third party coding tools, model routers and cloud platforms.
#xAI #launches #Grok #stronger #coding #knowledge #work #performance






