🚨Muse Spark 1.2 Leaderboard Rankings
I’ve added it across the main LuminaBench categories here are the main ones:
• Overall: 78.5
• Coding: 77.0
• Reasoning: 77.0
• Agents: 79.9
Breakdowns for each below👇
https://t.co/Ibhjm4bGvp
@tetsuoai Next year sometime they definitely will, but openAI have like two models ahead of what they already have released just waiting / near ready + astra will be 10T and they have something bigger still almost ready
@FReza1984 yeah especially with the progress they are making, it will get quantised and optimised and will run on a 16GB laptop at some point comfortably, not fast but good enough
🚨Qwen3.8 27B drops next Week
This might actually be a more interesting Qwen release than the 2.4T Max model:
• 27B parameters
• Open weights
• Drops August 12th
• This should also run locally on around 17GB RAM/VRAM when quantised, so a very low bar to entry
The 20 to 30B Parameter model range is getting very good, what you can actually run locally now compared to even 6 months ago is a huge jump.
Really excited for this one.
🚨Astra leaks are getting ridiculous
Quick correction first: Doug is apparently the codename for Astra’s pre train not a new one.
Latest claims:
• Astra will be around 10T parameters
• Built from the new Doug pre train
• Supposedly it crushes Fable across benchmarks👀
• Apparently a huge jump in writing quality from it too
• Still expected this month even with the setback
• And there’s an even bigger pre train leaked for the end of the year
I’m expecting this to be the first model that genuinely moves past Fable/Mythos and is a definitive leader across most benchmarks.
https://t.co/S84mEjjWJ5
🚨GPT 6 might not even be OpenAI’s biggest model this year
A much larger pre train is leaked to be coming around November time:
• It is Codenamed “Doug ”
• I think this is the model OpenAI alluded to in June
• This is also going to be OpenAI’s biggest pre train yet
• It is also supposedly strong enough to make Fable seem “primitive”😭
I genuinely can't fathom what this model will be capable of if Fable is primitive to it, this thing is going to be ridiculous.
https://t.co/Qzc4vKxWNa
@Giqnn6 This is genuine leaks from people who spoke to employees but of course, people love to know what is going to happen and speculate it’s half of the fun
🚨GPT 6 might not even be OpenAI’s biggest model this year
A much larger pre train is leaked to be coming around November time:
• It is Codenamed “Doug ”
• I think this is the model OpenAI alluded to in June
• This is also going to be OpenAI’s biggest pre train yet
• It is also supposedly strong enough to make Fable seem “primitive”😭
I genuinely can't fathom what this model will be capable of if Fable is primitive to it, this thing is going to be ridiculous.
https://t.co/Qzc4vKxWNa
GPT-5.5 will not be the last major pre-training run from OpenAI.
GPT-6 will be a great model. However, the end-of-year model I alluded to back in June is going to be OpenAI’s biggest pre-train, as far as I know.
Now we know that model is codenamed ‘Doug.’
And it will make Fable seem ‘primitive.’
🚨A new OpenAI image model just surfaced on Arena
It’s showing up under the codename “mona-lisa-1”, and looks to be a new unreleased OpenAI image model.
This could be what OpenAI has been A/B testing over the past few weeks, and it looks like more than just a small upgrade.
https://t.co/sdFhUOrQ4R
🚨New OpenAI image model A/B tests spotted
• Could be a new model or a GPT Image 2 update
• Tests appear on prompt following and text rendering
• OpenAI used A/B tests before releasing GPT Image 2
GPT Image 2 is already the best Image model, this would put them even further in the lead.
Have you noticed any of these A/B tests recently?