@ate_bites@AdamHoltererer The numbers are meaningless. What does a score of 60 mean on the Intelligence Index versus 50? The 60 model might be worth ten times more if it can do something the 50 model can't do.
@tukangoprekk@claudeai So far it's meant that when you use Claude Fable, it counts two ticks towards a Fable-specific limit for every tick that it counts against the total weekly limit.
@Apollo_EO@synthwavedd OpenAI isn't giving us their best either. GPT-5.6 Sol High is very different from GPT-5.6 Sol Max. And AA briefcase says that on pure rubric store, even Sonnet 5 Max outperforms GPT-5.6 Sol Max. I think Sol is good at narrow tasks, but misses things along the way.
@ArtificialAnlys I think it's great that you guys split it, but that means combined the ELO overstates GPT's performance on analytical quality then!
Classic GPT.