🎮 We introduce GameHorizon Suite, a large-scale data and evaluation suite spanning multiple horizons and AAA games. It serves as a standardized yardstick for VLMs, UMMs, GUI, coding, and game agents.
Github: https://t.co/Q3MS8A53eT
HF Daily: https://t.co/GMw7gJ7wl0
Project Page: https://t.co/bN5JpvDWsA
Introducing GameHorizon: a unified benchmark for AI gameplay across games, model families and task scales.
• Challenge: Existing benchmarks often cover few games, lack language instructions or rely on high-variance online rollouts, making consistent model comparison difficult.
• What: A unified data and evaluation suite spanning multiple temporal horizons, diverse AAA games and a broad range of model families.
• How: Aligning videos and actions with automatically annotated operations, goals and strategies—combined with reproducible offline evaluation and stepwise online testing.
• Result: Evaluation of 47 models reveals that future-action planning and goal decomposition remain key bottlenecks. Among the 12 models also tested online, stronger offline performance generally corresponds to better gameplay results.
Tencent ARC Lab just released GameHorizon Suite
A unified data and evaluation suite measuring AAA gameplay capabilities across multiple temporal horizons for VLMs, UMMs, GUI, coding, and game agents.
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?
I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev
• 20-200x faster
• 40-400x cheaper (w/ output tokens free)
• Frontier composable intelligence optimized for decisions
AFAICT the shortest path to AI-based economic revolution
Today is a very historical moment for AI video generation
You can now generate AI video faster than you can watch it
Before it'd take let's say 2-5 minutes to generate 15 seconds of video
@fal made a post-trained Minimax H3 variant called Max which is 50x faster than the original but still maintains quality
It generates 15 seconds of video in 9 seconds!
That means you can now do new things like build a perpetual livestream with it that never ends!
Introduce #EditaLive, our new model for streaming character animation with instruction-based Editing (single forward!!!). Arxiv: https://t.co/dU3Z28rGXV; Code: https://t.co/g2cLsxS7CX
We’re reimagining a 50-year-old interface - the mouse pointer - with AI. 🖱️
These experimental demos show how people can intuitively direct Gemini on their screens using motion, speech, and natural shorthand to get things done 🧵
CutClaw: Agentic Hours-Long Video Editing via Music Synchronization
An autonomous multi-agent framework that transforms hours of raw footage into cinematic montages. It leverages MLLMs as a Playwriter, Editor, and Reviewer to orchestrate storytelling and synchronize cuts with music beats.
This AI allows you to stream as anyone in realtime.
Free & open source, works with regular GPU.
It was a pain to install but I eventually got it to work. Full tutorial: https://t.co/CiKdjY8F7B
Cute RL paper from Disney Research for the holidays❄️
"The illusion of believability is fragile: even small inconsistencies, such as rough foot impacts or jitter, can break the character's lifelike appearance."
To maximize the character's believability, they used RL to optimize for realism & acoustic stealth.
They basically conditioned the policy directly on artist-created animation files, training the agent to learn Olaf's specific rolling gait from the movie rather than a standard efficient robot walk.
To solve the "robotic clanking" problem, they introduced an "impact reduction" reward that penalizes high vertical velocity changes specifically at the moment of foot contact, effectively teaching the robot to "tiptoe".
Happy holidays / Merry Christmas!🎄❄️
PersonaLive! Expressive Portrait Animation
A novel diffusion framework by GVC Lab & University of Macau researchers enables real-time, streaming, infinite-length portrait animations with 7-22x speedup on a single 12GB GPU.
Wan again? No, it's SD! PersonaLive!: Your vtuber dreams are now real-time, streamable & infinite-length;
- animates any portrait at 15.82 FPS with just 0.253s latency
https://t.co/qgseSefV7Z