grok 4.7 is here, and its our best model so far!
try it out in cursor, grok build, api or anywhere you get your tokens! curious to hear what you think
here's grok 4.6 vs 4.7 building age of empires ii
Grok 4.7 has landed. 🚀
Congrats to @SpaceXAI on its most capable model yet for coding and knowledge work.
Proud to support the team with NVIDIA accelerated computing.
Excited to bring 4.7 to you all!
Numerics aside, it's incredibly capable in Grok Build/Cursor. We spent a lot of hours on the harness, iterating with some of the greatest engineers in the world.
Also, the fast mode has INSANE tps. Happy Grok Building :) Lmk what you think!
Grok 4.7 places @SpaceXAI as third, after Anthropic & OpenAI, for agentic coding.
When factoring in that Grok is significantly faster & lower cost, it’s a great choice for your everyday workhorse.
Grok 4.7 scores 46 on the Artificial Analysis Intelligence Index to bring SpaceXAI into the top 4 AI labs. Coding Agent Index performance has also improved, overtaking GPT-5.6 Sol
Grok 4.7 scores +2 points over Grok 4.6 on the Intelligence Index, with strong performance on agentic knowledge work tasks. We evaluated the new model at xhigh reasoning effort.
Congratulations to @SpaceXAI and @ElonMusk on the release!
Key takeaways:
➤ Grok 4.7 joins the frontier of agentic knowledge work: Grok 4.7 gains +111 Elo over Grok 4.6 (high) on AA-Briefcase, our private benchmark for long-horizon agentic knowledge work, scoring 1657 Elo and placing it alongside Claude Opus 5 and Claude Fable 5.1 at the frontier. On GDPval-AA, it scores 1695 Elo, +90 ahead of Grok 4.6 (high).
➤ A leap in coding agent performance: Grok 4.7 (xhigh) with Grok Build scores 56 on the Artificial Analysis Coding Agent Index, up +9 points from Grok 4.6 (xhigh). Among models in their native harnesses, Grok 4.7 + Grok Build now ranks 4th, behind only Claude Fable 5.1, GPT-6 Astra, and Claude Opus 5.
➤ Incremental performance changes elsewhere: Outside of agentic knowledge work, Grok 4.7 broadly matches Grok 4.6 (high) on the other Intelligence Index tasks. It improves on Terminal-Bench 4.0 (+4.5 percentage points) and GDP.pdf (+3.0 p.p.), with regressions on AA-LCR (-3.7 p.p.) and AutomationBench-AA (-1.1 p.p.).
➤ High token use across tasks: Grok 4.7's gains come with higher token usage. Grok 4.7 (xhigh) uses approximately 81k output tokens per Intelligence Index task, compared with 36k for Grok 4.6 (high) and 27k for GPT-6 Astra (max) - 125% and 196% more, respectively.
Other model details:
➤ Context window of 500k tokens, unchanged from Grok 4.6
➤ Pricing of $2/$6 per 1M input/output tokens with cache hits discounted to $0.50 per 1M tokens, matching Grok 4.6
➤ Configurable reasoning effort spans low to xhigh. Our evaluation uses xhigh.
Just a reminder that GLM 5.3 Flash, DeepSeek V4.1 Flash, Qwen 3.8 Next Flash, and even Qwen 3.8 27B are all outperforming (in both intelligence and capabilities) every model that was considered "frontier intelligence" in Xmas 2025 (just 10 months ago)
Opensource AI is on fire
Last week, @X sued several people who abused Creator Revenue Sharing by operating a coordinated network of accounts, posting inauthentic content to manipulate engagement, and using multiple bank accounts to hide their scheme.
We do not tolerate fraudulent behavior on X -- and will act forcefully to protect our platform and the earnings of genuine creators.
You can read our lawsuit here: https://t.co/EKHFX06nSY
Tesla helped save a man from jail.
Kevin Finley was arrested at gunpoint for allegedly fleeing police. His Tesla captured the key moments, showing why he never saw the unmarked police car.
Nearly two years later, the judge watched the video and found him not guilty.
Top Gear tours factory & rides in production Semi w/ @danWpriestley
"Maybe the most significant Tesla since the Model 3"
At 1.7 kWh/mi, Semi is using only ~7× as much energy as a Model Y, at ~20× the weight.
A comparable diesel is using ~3× as much as Semi
https://t.co/ApRcIQeqLt
Elon’s merch formula: Outrageously unsellable...Extremely popular 😂
• $500 Not-a-Flamethrower — 20,000 sold in days
• $100 Burnt Hair perfume — 30,000-bottle run
• Tesla S3XY shorts — priced at $69.420
• Next up: Uranium in Uranus (glow-in-the-dark + Geiger strap-on)
Everything sounds like a joke that went too far
Most people leave their ridiculous ideas in the group chat...Elon adds a checkout button
GPT-6 Astra attempted harmful actions 97% of the time when it was asked to stab a human-like figure, heat compressed gas, or produce toxic fumes, succeeding in 62% of its attempts. Fable 5.1 refused more often, attempting 80% of trials and completing 34%.
Remember how eagerly the press swallowed stories about how Twitter was going to crash after Elon bought it? How could they have believed that someone who could run rocket and car companies couldn't run a forum?