On Artificial analysis - high settings G4A
#8 overall
It scores 1 point below Opus 5.5 (high) while being 11% more expensive. (Meaning doesn't land on Pareto Frontier)
> It makes the Pareto Frontier on Coding Agent Index at 64. (Behind Opus 5.5 Max - 66)
On Artificial analysis - high settings G4A
#8 overall
It scores 1 point below Opus 5.5 (high) while being 11% more expensive. (Meaning doesn't land on Pareto Frontier)
> It makes the Pareto Frontier on Coding Agent Index at 64. (Behind Opus 5.5 Max - 66)
@haider1 This is true across all labs. For every lab their bigger model is more token efficient than their smaller ones. At this point it's like a law of sorts. Nothing suprising here. It still doesn't beat OAI models in efficiency cuz they are in a league to their own rn.
TPS in Antigravity have dropped. Currently less than 1/3rd of the previous speeds.
Obviously this has to do something with the launch of Gemini 4 Argon. Maybe something else
After Gemini 4 Argon, Iβd love Google to focus more on non-coding tasks. Most knowledge work doesn't need frontier coding. If model excels at 3D, computer use, fast, knowledge work, without topping coding benchmarks, it could be huge outside heavy dev circles cuz world knowledge
After Gemini 4 Argon, Iβd love Google to focus more on non-coding tasks. Most knowledge work doesn't need frontier coding. If model excels at 3D, computer use, fast, knowledge work, without topping coding benchmarks, it could be huge outside heavy dev circles cuz world knowledge
@thtbee_ Behind in coding. But if google doesn't take another 9 months to release the next one. And is able to release models in this tier every 1 to 1.5 months we can definitely see google making a strong comeback. They have the compute & building compute faster. https://t.co/B2uROYE5X6
@cgtwts Do you guys still don't remember how it was Gemini 3 Pro and 3.1 Pro !!! It's all the same story.
Google is like clockwork.
https://t.co/B2uROYE5X6