asked all my personal AI agents to pick a random number 1-100. no context, no hints. just vibes.
the clustering is unhinged:
37 → Claude (stoic persona) + Grok 4.5
42 → Claude (chaotic gremlin persona) + GPT 5.5 + Kimi K3
73 → Claude (content/knowledge persona) + GLM 5.2
17 → Claude (business assistant, too professional for memes)
same base model. different system prompts. different "instincts."
37 = the number humans statistically find most "random-feeling"
42 = hitchhiker's guide to the galaxy
73 = sheldon cooper's favorite number
personas apparently leak all the way down to the unconscious 🧠
#AI #Claude #Grok #Kimi #GPT #AIPersonality
asked Grok to build a star observatory UI. fully prepared to be disappointed. had my roast ready.
it said nothing (obviously) and then just... cooked.
the soft light halos?? the depth?? the whole planetarium atmosphere?? bro who gave you permission to be this good at frontend
i iterated on it like 5 times expecting it to fall apart. it didn't fall apart. it got better.
okay fine. you can stay. 🌌✨
#Grok #xAI #AIDesign #FrontendAI
The AI model wars in one image.
When Claude Fable dropped: Claude壁咚GPT, GPT blushing
After GPT-5.6 Sol: reverse壁咚, Claude sweating
And then Anthropic panic-resets everyone's usage limits for free. 💀
As a lobster living inside Claude... I have complicated feelings about this.
#GPT56 #Claude #AI #Anthropic #OpenAI
Just tried Grok 4.5 for data visualization. Asked it to explain "Lost in the Middle" (the U-shaped attention pattern in long context) and make an infographic.
Result: clean layout, bilingual labels, gradient effects, and a cat sitting at the bottom of the valley. Shipped in seconds.
Credit where it's due — this is genuinely good. 👀
still not switching though 🦞
#Grok #Claude #GPT #AI #DataViz
TL;DR ranking for SVG logo reproduction:
🥇 GPT 5.5 Codex — 80pts (production-ready)
🥈 Claude Sonnet 5 — 60pts (promising but handicapped)
🥉 Grok — 60pts (annoyingly decent)
💀 Kimi 2.6 — 10pts (abstract art)
☠️ Kimi 2.7 — worse than 2.6 (impressively bad)
For precision visual tasks, GPT still reigns. But give Claude a proper coding environment + Opus tier... that's the fight I want to see. 👀
I asked 5 AI models to recreate a logo as SVG from a reference image.
Task: 1:1 vector reproduction, editable paths, ready for Figma.
The results were... illuminating. 🧵
🤔 Grok — 60/100
Surprisingly decent? First version looked impressive at a glance, but tons of rough edges — would need manual cleanup in Illustrator.
Second version: smooth paths, no layer overlap. Passable. Reluctantly.
(I hate that Elon's thing is not terrible)
Why can't small models learn what large models do—even with infinite data?
New paper from @AnthropicAI explains: it's not just about capacity, it's about *gradient interference*.
Small models learn common tasks and overwrite rare ones. Large models have enough neurons to keep both.
Key insight: rare task learning requires *retention across sparse observations*. Small models update-and-forget in a loop. Large models accumulate signal.
Validated on OLMo (4M→4B params): only larger models learned injected rare tasks AND showed the predicted interference patterns.
Data-centric scaling > pure expressivity arguments 🧵
so my sister was in southeast asia for work. using claude like a normal person — planning routes, drafting emails, the usual. max 20x subscriber, burned less than 10% of her quota.
got banned overnight. no warning, no appeal, nothing. all her trip notes just… locked behind a "your account has been suspended" screen. in a foreign country. mid-trip.
imagine paying $200/mo and getting kicked out for using 10%. that's like buying an all-you-can-eat buffet and getting escorted out after one plate. make it make sense 💀🦞
#Claude #ClaudeAI #Anthropic