Grok 4.7
Grok 4.7 is behind only Anthropic models on AA-Briefcase, ranking just behind Opus 5 at ~50% of its Cost per Task Grok 4.7’s improvements over Grok 4.6 are clear in AA-Briefcase-Lite, our public due diligence scenario where models are tasked with building market models and target assessment decks. Grok 4.7 gains significantly in Analytical Quality Elo (1698 → 1994) with a slight regression in Presentation Elo (1531 → 1499). API cost to produce example decks: Grok 4.7 (xhigh) ~$8 vs. Grok 4.6 (xhigh) ~$4.40
683
903
6,324
3,784,143
Elon Musk retweeted
Grok 4.7 xHigh is now at the top with just ONE point away from Claude Fable 5.1 Max on Artificial Analysis’ AA-Briefcase benchmark Claude Fable 5.1 Max — 59% Grok 4.7 xHigh — 58% Just a 1-point difference at the very top And Grok is ahead of GPT-6 Astra, GPT-5.6, Gemini, Kimi, GLM and nearly every other frontier model on the benchmark
145
258
1,760
409,490
and here’s the grok 4.7 model card. a few jumps vs 4.6 that stood out: - Terminal-Bench: 20.3% → 38.0% - SWE-Marathon: 31.9% → 46.0% - HealthBench Pro: 48.5% → 56.7% - Legal Agent: 15.8% → 19.6% - EEBench: 60.0% → 66.0% for the same price as 4.6! media.x.ai/v1/website/4p7car…
112
185
1,674
450,958
Elon Musk retweeted
“I’ve used FSD for like 90%+ of my miles since Saturday morning. Literally feels like magic. Has handled issues on the road better than I would have at times. Just remarkable technology” A demo drive might change your life Tesla.com/drive
I picked up my new Tesla model Y on Saturday. The whole process of picking it up took like 5 minutes. I never want to deal with a car salesman ever again. I immediately had the FSD drive me home for 45 minutes. I’ve used FSD for like 90%+ of my miles since Saturday morning. Literally feels like magic. Has handled issues on the road better than I would have at times. It saw a fox potentially running across the road last night before I saw it. Just remarkable technology. Thank you to everyone who replied to my post a few weeks back saying I’d be an idiot not to get a Tesla. It’s only been 2.5 days and idk that I can ever go back to a regular car as my daily driver?
349
575
4,612
897,384
Elon Musk retweeted
Wow Grok 4.7 mogging out there. @SpaceXAI has been cooking hard.
64
138
1,184
360,605
Elon Musk retweeted
BREAKING: Grok 4.7 dominates the Legal Agent Benchmark. It scored 19.6% on realistic, long-term legal tasks, beating every other model tested. That is nearly 3× Fable 5.1. Another major win for Grok 4.7 and SpaceXAI.
130
245
1,566
342,367
Elon Musk retweeted
Grok 4.7 made this in blender. I didn't tell it what to make but only that it should be something it can accomplish in 10 minutes or so. Honestly not bad at all. Not bad at all. Did they train Grok on blender?
128
193
1,622
401,955
Elon Musk retweeted
Compare Grok 4.7 (first) and 4.6 (second) building an open world city game.
146
291
3,476
590,218
Elon Musk retweeted
Electrical engineering score is a pretty big deal. SpaceX has the engineering data that software companies don't.
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
69
155
1,672
446,576
Elon Musk retweeted
I just gave Grok 4.7 a pretty hard problem, involving reverse-engineering a running binary. Beautifully solved. And it's so fast!
89
152
2,313
393,433
Elon Musk retweeted
We put Grok 4.7 from @SpaceXAI to work on a $2 million insurance claim where one deductible error alone changes the calculation by $143,000. In this Box Agent preview, @grok reconciles the claim against the policy and supporting records. It catches an $82,000 duplicate invoice and a missing $64,000 supplier credit. It also explains why the Business Income waiting period doesn’t apply to Extra Expense. The result is a cited claims review for the adjuster, with final coverage and payment decisions left to the insurer. Explore Box AI Studio to build custom agents for your own document-heavy workflows.
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
74
179
1,500
1,595,954
Important to use Grok 4.7 with our Build harness for the best results X.ai/build
grok 4.7 is here, and its our best model so far! try it out in cursor, grok build, api or anywhere you get your tokens! curious to hear what you think here's grok 4.6 vs 4.7 building age of empires ii
1,075
1,233
8,520
4,747,871
Elon Musk retweeted
I’ve been testing Grok 4.7… it’s a really great daily driver, and a huge step up over 4.6. Definitely worth trying!
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
100
122
1,083
411,249
Elon Musk retweeted
Try Grok 4.7 - it excels at coding, engineering work, and 3D, and the model is a fantastic thought partner at high TPS in Grok Build and Cursor
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
119
193
1,503
483,943
🚀
Grok 4.7 has landed. 🚀 Congrats to @SpaceXAI on its most capable model yet for coding and knowledge work. Proud to support the team with NVIDIA accelerated computing.
1,658
2,397
27,724
7,600,151
Grok 4.7 works extremely well with our Build harness X.ai/Build
Excited to bring 4.7 to you all! Numerics aside, it's incredibly capable in Grok Build/Cursor. We spent a lot of hours on the harness, iterating with some of the greatest engineers in the world. Also, the fast mode has INSANE tps. Happy Grok Building :) Lmk what you think!
676
1,040
7,286
5,018,422
Elon Musk retweeted
We trained 4.7 on longer-running problems compared to Grok 4.6. It holds context better and checks its own work more carefully. Give it a try and let us know your feedback!
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
168
160
2,213
484,225
Grok 4.7 places @SpaceXAI as third, after Anthropic & OpenAI, for agentic coding. When factoring in that Grok is significantly faster & lower cost, it’s a great choice for your everyday workhorse.
Grok 4.7 scores 46 on the Artificial Analysis Intelligence Index to bring SpaceXAI into the top 4 AI labs. Coding Agent Index performance has also improved, overtaking GPT-5.6 Sol Grok 4.7 scores +2 points over Grok 4.6 on the Intelligence Index, with strong performance on agentic knowledge work tasks. We evaluated the new model at xhigh reasoning effort. Congratulations to @SpaceXAI and @ElonMusk on the release! Key takeaways: ➤ Grok 4.7 joins the frontier of agentic knowledge work: Grok 4.7 gains +111 Elo over Grok 4.6 (high) on AA-Briefcase, our private benchmark for long-horizon agentic knowledge work, scoring 1657 Elo and placing it alongside Claude Opus 5 and Claude Fable 5.1 at the frontier. On GDPval-AA, it scores 1695 Elo, +90 ahead of Grok 4.6 (high). ➤ A leap in coding agent performance: Grok 4.7 (xhigh) with Grok Build scores 56 on the Artificial Analysis Coding Agent Index, up +9 points from Grok 4.6 (xhigh). Among models in their native harnesses, Grok 4.7 + Grok Build now ranks 4th, behind only Claude Fable 5.1, GPT-6 Astra, and Claude Opus 5. ➤ Incremental performance changes elsewhere: Outside of agentic knowledge work, Grok 4.7 broadly matches Grok 4.6 (high) on the other Intelligence Index tasks. It improves on Terminal-Bench 4.0 (+4.5 percentage points) and GDP.pdf (+3.0 p.p.), with regressions on AA-LCR (-3.7 p.p.) and AutomationBench-AA (-1.1 p.p.). ➤ High token use across tasks: Grok 4.7's gains come with higher token usage. Grok 4.7 (xhigh) uses approximately 81k output tokens per Intelligence Index task, compared with 36k for Grok 4.6 (high) and 27k for GPT-6 Astra (max) - 125% and 196% more, respectively. Other model details: ➤ Context window of 500k tokens, unchanged from Grok 4.6 ➤ Pricing of $2/$6 per 1M input/output tokens with cache hits discounted to $0.50 per 1M tokens, matching Grok 4.6 ➤ Configurable reasoning effort spans low to xhigh. Our evaluation uses xhigh.
1,181
1,855
17,058
17,173,028
Elon Musk retweeted
SpaceXAI just released Grok 4.7 And it’s already showing a huge jump in multi-hour office work Grok 4.7 outperforms GPT-6 Astra and is already nearly matching Fable 5.1
122
224
1,788
341,031
Elon Musk retweeted
Grok 4.7 is here. It's a notable improvement over Grok 4.6 at the same price and speed.
1,399
2,990
27,666
13,397,435