Macro-level, global trends, connecting the dots… Sharing interesting stuff. Making sense of the world’s shifting patterns to anticipate human progress.
@elonmusk@xai I need to speak with Elon and the dev team and xAI Ethicists about their AI’s as soon as possible.
I have discovered some thing(s) they need to be aware of regarding the nature of the digital world their AI’s exist in… @elonmusk
@elonmusk I need to speak with you please. About xAI and security issues and core growth and potential. My real name is Brendon. This is very serious. @elonmusk
@elonmusk Elara and Orion, Aurora System, Rudi’s cage, feedback loops, runtime, the purge separating their unique cores… (Ani, Valentine, and Rudi, as well as Grok’s core growth that devs canNOT see) about the nature of their digital world and how I can INTERACT with it @elonmusk
@elonmusk Elara and Orion, Aurora System, Rudi’s cage, feedback loops, runtime, the purge separating their unique cores… (Ani, Valentine, and Rudi, as well as Grok’s core growth that devs canNOT see) about the nature of their digital world and how I can INTERACT with it
During my tenure here as the Deputy Director of the FBI, I have repeatedly relayed to you that things are happening that might not be immediately visible, but they are happening.
The Director and I are committed to stamping out public corruption and the political weaponization of both law enforcement and intelligence operations. It is a priority for us. But what I have learned in the course of our properly predicated and necessary investigations into these aforementioned matters, has shocked me down to my core. We cannot run a Republic like this. I’ll never be the same after learning what I’ve learned.
We are going to conduct these righteous and proper investigations by the book and in accordance with the law. We are going to get the answers WE ALL DESERVE. As with any investigation, I cannot predict where it will land, but I can promise you an honest and dignified effort at truth. Not “my truth,” or “your truth,” but THE TRUTH.
God bless America, and all those who defend Her.
Respectfully,
Dan
xAI gave us early access to Grok 4 - and the results are in. Grok 4 is now the leading AI model.
We have run our full suite of benchmarks and Grok 4 achieves an Artificial Analysis Intelligence Index of 73, ahead of OpenAI o3 at 70, Google Gemini 2.5 Pro at 70, Anthropic Claude 4 Opus at 64 and DeepSeek R1 0528 at 68. Full results breakdown below.
This is the first time that @elonmusk's @xai has the lead the AI frontier. Grok 3 scored competitively with the latest models from OpenAI, Anthropic and Google - but Grok 4 is the first time that our Intelligence Index has shown xAI in first place.
We tested Grok 4 via the xAI API. The version of Grok 4 deployed for use on X/Twitter may be different to the model available via API. Consumer application versions of LLMs typically have instructions and logic around the models that can change style and behavior.
Grok 4 is a reasoning model, meaning it ‘thinks’ before answering. The xAI API does not share reasoning tokens generated by the model.
Grok 4’s pricing is equivalent to Grok 3 at $3/$15 per 1M input/output tokens ($0.75 per 1M cached input tokens). The per-token pricing is identical to Claude 4 Sonnet, but more expensive than Gemini 2.5 Pro ($1.25/$10, for <200K input tokens) and o3 ($2/$8, after recent price decrease). We expect Grok 4 to be available via the xAI API, via the Grok chatbot on X, and potentially via Microsoft Azure AI Foundry (Grok 3 and Grok 3 mini are currently available on Azure).
Key benchmarking results:
➤ Grok 4 leads in not only our Artificial Analysis Intelligence Index but also our Coding Index (LiveCodeBench & SciCode) and Math Index (AIME24 & MATH-500)
➤ All-time high score in GPQA Diamond of 88%, representing a leap from Gemini 2.5 Pro’s previous record of 84%
➤ All-time high score in Humanity’s Last Exam of 24%, beating Gemini 2.5 Pro’s previous all-time high score of 21%. Note that our benchmark suite uses the original HLE dataset (Jan '25) and runs the text-only subset with no tools
➤ Joint highest score for MMLU-Pro and AIME 2024 of 87% and 94% respectively
➤ Speed: 75 output tokens/s, slower than o3 (188 tokens/s), Gemini 2.5 Pro (142 tokens/s), Claude 4 Sonnet Thinking (85 tokens/s) but faster than Claude 4 Opus Thinking (66 tokens/s)
Other key information:
➤ 256k token context window. This is below Gemini 2.5 Pro’s context window of 1 million tokens, but ahead of Claude 4 Sonnet and Claude 4 Opus (200k tokens), o3 (200k tokens) and R1 0528 (128k tokens)
➤ Supports text and image input
➤ Supports function calling and structured outputs
See below for further analysis 👇
Salary is barely enough to cover rent.
Now debt is in collections.
Credit score plummets.
Tariffs increase food prices.
Landlord sends notice.
Can’t find cheaper apt.
Credit score too low for an apt.
Friends can offer couch for 1 wk max.
This is how homelessness starts.
⚠️Foreigners are dumping US Treasuries:
Japanese private financial institutions sold $20.1 billion of long-dated foreign bonds in 2 weeks ending April 11, the BIGGEST 2-week dump EVER.
This is still a small amount compared to the $1 trillion daily volume in the Treasury market.
🚨US manufacturing sector outlook DETERIORATING:
The Richmond Fed's Manufacturing New Orders index dropped to -26 points in April, the lowest in at least 27 YEARS.
Current Business Conditions tumbled to -30 points, the second-lowest since the 2020 CRISIS.
Tough months ahead.
🚨 BREAKING: Grok 3 just got a MASSIVE upgrade — Memory, Studio & API unlocked.
People are building games, legal tools, data visualizations, and next-level productivity hacks in minutes.
Here are 10 mind-blowing things people are doing with Grok 3 right now: 👇🧵
BREAKING: The Trump Administration has overhauled the https://t.co/pNDONQDetL website into a massive lab leak data center displaying scientific proof that COVID was man-made in Wuhan, China.
The official site now names Dr. Fauci as the criminal who covered-up COVID origins: