Designer, photographer, and visual artist / he/his/him / All content and viewpoints expressed herein are my own and not those of my current or past employers.
🚨BREAKING: OpenAI published a paper proving that ChatGPT will always make things up.
Not sometimes. Not until the next update. Always. They proved it with math.
Even with perfect training data and unlimited computing power, AI models will still confidently tell you things that are completely false. This isn't a bug they're working on. It's baked into how these systems work at a fundamental level.
And their own numbers are brutal. OpenAI's o1 reasoning model hallucinates 16% of the time. Their newer o3 model? 33%. Their newest o4-mini? 48%. Nearly half of what their most recent model tells you could be fabricated. The "smarter" models are actually getting worse at telling the truth.
Here's why it can't be fixed. Language models work by predicting the next word based on probability. When they hit something uncertain, they don't pause. They don't flag it. They guess. And they guess with complete confidence, because that's exactly what they were trained to do.
The researchers looked at the 10 biggest AI benchmarks used to measure how good these models are. 9 out of 10 give the same score for saying "I don't know" as for giving a completely wrong answer: zero points. The entire testing system literally punishes honesty and rewards guessing.
So the AI learned the optimal strategy: always guess. Never admit uncertainty. Sound confident even when you're making it up.
OpenAI's proposed fix? Have ChatGPT say "I don't know" when it's unsure. Their own math shows this would mean roughly 30% of your questions get no answer. Imagine asking ChatGPT something three times out of ten and getting "I'm not confident enough to respond." Users would leave overnight. So the fix exists, but it would kill the product.
This isn't just OpenAI's problem. DeepMind and Tsinghua University independently reached the same conclusion. Three of the world's top AI labs, working separately, all agree: this is permanent.
Every time ChatGPT gives you an answer, ask yourself: is this real, or is it just a confident guess?
On the first day of Black History Month, let me share some scathing words from the great Fredrick Douglass regarding the American church of his time. I leave it to you to decide for yourself as to whether his criticisms still stand today.
“Officers were asking everyone what country they were from, and if they said a certain country, they were told to step out of line and that their oath ceremonies were canceled.”
https://t.co/gYLYnaxkmy
Here at Galadriel, we are building a different kind of AI.
Instead of a Dark Lord, it will be a queen. Not dark but beautiful and terrible as the dawn! Tempestuous as the sea, and stronger than the foundations of the earth! All shall love it and despair!
Happy Halloween, y’all! I get to spend it measuring a pitch-black bank vault in the basement of a weird granite and marble Victorian-era Egyptian-Mesopotamian-eclectic revival building in New Orleans. What could go wrong?
The majority of the No Kings protests have dispersed at this time and all traffic closures have been lifted.
We had more than 100,000 people across all five boroughs peacefully exercising their first amendment rights and the NYPD made zero protest-related arrests.
NYPD announces that zero arrests were made at the massive No Kings protests today after Republicans spent days claiming the protestors are violent terrorists.