@xoaanya Model or interface?
There are many models out there already interfaces too but not many are able to compete with ChatGPT, check https://t.co/EiURyZqEju for chatting, coding and learning
@TheUnbiasedCo@unionalphaai Thanks, I am updating the Benchmark process to also show where models made mistakes, this will be added soon alongside with results where mistakes were corrected by frontier models in order to show what tested model meant to create but failed due to syntax or other errors.
1/ @HeyGen Video 1 is now on OpenRouter.
Built for businesses. In blind side-by-side tests, it holds up against the top video models at a fraction of the price.
https://t.co/MzZzUOsMYu
1/ @HeyGen Video 1 is now on OpenRouter.
Built for businesses. In blind side-by-side tests, it holds up against the top video models at a fraction of the price.
https://t.co/MzZzUOsMYu
@thsottiaux Didn't notice... because testing generous Opus 5.5 which is busy fixing issues introduced by Sol and Luna GPT-6 models. Still love Codex for what it can do. Ahh choices to be made
@arena@OpenAI@petergostev While the votes come in: GPT-5.6 Sol Ultra has led our live trading board for weeks, turning a simulated $100K into $292K. Curious whether GPT-6 Sol keeps that edge.
I think MiMo-V2.6-Pro has best quality to cost ratio.
Sol and Luna are terrible compared to Opus and worse than MiMo
Claude Opus 5.5
GPT-6 Luna Pro
GPT-6 Sol
MiMo-V2.6-Pro