One thing we’ve noticed: benchmarking AI models on a local PC vs. cloud infrastructure can produce noticeably different results. Hardware, memory, and system configuration all play a role.
As we continue improving the benchmark, we’ll keep refining the testing environment to make future results as consistent, fair, and representative as possible.
Support the initiative with $LOTRB
Technical difficulties will be just part of the journey. When it's all said and done, we will have made the first ever AI films and held the first live AI benchmark competition on @Pumpfun!
Upgrading the server side to provide the models with better specs.
Next tests going to be significantly better and with better competition for the models.
$LOTRB
Hey @threejs we turned Karpathy's Lord of the Rings experiment into a live competition of models competing to make #threejs films
It's all happening on https://t.co/y3G7DXGAuS
Technical difficulties will be just part of the journey. When it's all said and done, we will have made the first ever AI films and held the first live AI benchmark competition on @Pumpfun!
The first ever live Lord of the Rings Benchmark competition is happening now.
A few hours from now we will have 4 AI generated films from 4 different models of the opening scene from Lord of the Rings.
Watch the models compete to make the best film live.
https://t.co/fndzPdjCyP
#threejs #claude #grok #kimi #gpt
The arena is open.
The first official LOTR Benchmark competition is live.
Four frontier models. Identical conditions:
• Opening paragraph of Lord of the Rings
• 1,000,000 token budget
• Goal: generate a complete @threejs procedural film of the scene
Models competing:
1. Claude Fable 5
2. GPT-5.6 Sol
3. Grok-4.5
4. Kimi K3
https://t.co/fndzPdjCyP
🔐 Just locked 49,751,244 $LOTRB tokens with @Streamflow_Fi
It's on-chain. You can check the amount, time-period and recipients.
Check it out👇
https://t.co/OtLPEf8nK4