Gemini 3.5 Pro is going to surprise a lot of people.
Google has not shipped a frontier model in 6 MONTHS.
Everyone is counting them out.
Look at the chart instead. Every green square is a Google release. They ship less often than any lab. And every single time it is a huge jump straight back to the frontier.
Silence, then leapfrog. The pattern has never broken.
If this pattern continues Gemini 3.5 Pro could outperform GPT 5.6 and Fable 5.
I am bullish on Gemini 3.5 Pro.
I still can’t figure out why Gemini struggles to compete with Claude and GPT.
- Owns Chrome
- Backed by Android
- Stores most search results
- Holds ~95% search history
- Google has the biggest user data
- Even incognito data isn’t fully private
So what’s the problem?
nvidia is casually giving you access to 5 frontier chinese AI models for free 😳
no credit card
no subscriptions
just one API key that unlocks everything
what you get for $0:
- DeepSeek V4 Flash for ultra-fast inference
- MiniMax M3 as a drop-in coding assistant
- Qwen3.5-397B for advanced reasoning tasks
- Kimi K2.6 for agentic workflows and long chains
- GLM 5.1 as a reliable everyday model
why this is huge:
> no paying separate subscriptions for different models
> no changing your existing workflows or tools
> no vendor lock-in since everything is OpenAI-compatible
getting started takes less than 2 minutes:
1. go to https://t.co/q50rSatNbb
2. sign up and verify your account
3. generate your nvapi key
4. set your base URL to https://t.co/92kkbFSf8D
5. pick any model and start building
supported models:
> minimaxai/minimax-m3
> qwen/qwen3.5-397b-a17b
> moonshotai/kimi-k2.6
> zhipuai/glm-5.1
> deepseek/deepseek-v4-flash
pro tip:
use DeepSeek V4 Flash for speed, Qwen for hard reasoning, Kimi for agents, and MiniMax as your daily coding companion
the best part?
one free key gives you access to 100+ models across NVIDIA's catalog
~40 requests per minute is more than enough for most developers and personal projects
5 frontier models that compete with GPT and Claude, all without spending a dollar
bookmark this and claim your free API key before the limits change 👀
FABLE 5 CAME BACK NERFED.
We re-ran the July 1st version of Claude Fable 5 on BridgeBench.
The results are brutal:
Debugging: 86.2 → 25.9
Refactoring: 73.6 → 38.4
Hallucination: 75.9 → 61.7
The new guardrails are kicking in on way too many tasks and falling back to Opus 4.8.
This is not the model that got banned.
Anthropic owes everyone an explanation.
Anthropic is literally becoming the AI villain:
- Not communicating updates people care about
- Limiting usage of the best model
- Making a big deal about small things (Sonnet 5)
OpenAI looks like the hero now:
- Lots of updates from the team constantly
- Lots of usage resets (free usage)
- Eating the cost of their best model
I don't run a trillion dollar company, but this seems so shortsighted to me. Anthropic was CRUSHING OpenAI six months ago, but all of these decisions are slowly catching up to them and hurting customer sentiment.
Obviously the Fable situation with the U.S. Government is not easy to handle, but you can come out of it looking like the hero. Anthropic is not doing that. Instead, they're going to limit the included plan usage and then immediately switch models or charge more?? I don't get it, someone explain the logic?
I was fully out on OpenAI. Today, I use Codex and GPT-5.5 more than Claude/Opus. I wouldn't have thought that was possible a few months ago. I was Claude-pilled HARD.
All of this is a great lesson for the rest of us:
1. Communicate often and honestly
2. Don't cover up important things with trivial things
3. Give a ton of value for the cost
There's still time for Anthropic to save this, but they're definitely not trending in the right direction.
Sonnet 5 goes straight into the garbage bin
> 1.2x more expensive than Opus 4.8 Max
> 2x more expensive than GPT-5.5-xhigh
> 5x more expensive than GLM-5.2
> 7x more expensive than Kimi-K2.6
> 57x more expensive than DeepSeek-V4-Pro
My entire AI stack is now Chinese 🇨🇳
87% cheaper. same revenue
swaps by task:
1. reasoning / backend brain
Opus 4.8 → Kimi K2.7
benchmark gap: ~8% · price: ~11x cheaper
2. code generation
GPT-5.5 → Qwen 3.7 Max
benchmark gap: ~18% · price: ~7x cheaper
3. agent loops + tool calling
Sonnet 4.7 → GLM 5.2
benchmark gap: ~3% · price: ~5x cheaper on input
4. cheap volume / bulk processing
GPT-5.5 mini → MiMo V2.5
benchmark gap: ~6% · price: ~12x cheaper
5. image generation
GPT-Image-2 → Wan 2.5
benchmark gap: ~5% · price: ~8x cheaper
6. video generation
Sora 2 → Kling 3.0
benchmark gap: roughly equal · price: ~6x cheaper
[ result after 30 days: ]
operating costs dropped 87%, output quality dropped 4% on average, revenue unchanged
the most important that these models will be not banned in a month and i can run them locally
nobody will steal my data and i can learn them as i need
full article drops tomorrow with:
> exact routing logic per task type
> the 2 cases where I still pay for American
> the migration playbook anyone can copy in a weekend
VERY IMPORTANT to get migrated now, while it's not too late
Anthropic just showed a 24-minute workshop on how to actually do prompts for Claude.
Taught by the people who built it.
Free. No registration. No paywall.
I've seen $300 courses that don't cover what they teach in the first 8 minutes.
Watch it and bookmark it now.
More people need to be talking about this! This is one of, if not the best vault, ever done in gymnastics history.
Mahdi Olfati - Yurchenko double back with a full twist.
So I looked up the case of bone cancer survivor Wong Qiu Yeu - she was treated in HKL, having flown from Sibu. Survived cancer & the treatment but came home to abusive father/stepmom.
Cops weren't helpful.
RIP Qiu Yeu. You deserved better.
https://t.co/0CJeKWG3XX