The interesting part of persistent World Models isn’t just continuity.
It’s that creators start thinking about atmosphere, composition, and spatial feeling instead of isolated frames. Feels much closer to crafting a place.
Built this w/ World Labs + GPT 5.5 + Three.js ☀️🍃🐱
Neural networks might speak English, but they think in shapes.
Understanding their rich *neural geometry* is key to understanding how they work – and to debugging and controlling them with precision.
Starting today, we’re releasing a series of posts on this research agenda. 🧵
We’re on the verge of interactive, real-time, photorealistic video generation with what are called World Models. These require a fair amount of expensive compute, but costs will come down over time. Today, these interactions are essentially real-time dreams. They lack the persistence and shared logic that turns a video into a multiplayer game. Or concert. Or classroom. Or the holodeck.
We believe the ultimate architecture for gaming (and the holodeck) is Roblox Reality. It’s a hybrid architecture that marries the structured data and logic of the Roblox Engine and Roblox Cloud with the generative power of Video World Models. The Roblox Engine provides the underlying synchronized ground truth—the score, physics, multiplayer sync, etc.—while our video model acts as a Super Upsampler to layer on photorealistic detail.
We believe this will ultimately remove barriers to high-fidelity creation, allowing a team of three people to build a narrative-driven, photorealistic masterpiece in a single week. This is an early look at turning solitary AI dreams into a social, playable reality.
https://t.co/9SYizn5TV6
We're extremely excited to announce the release of PATINA, a first-of-its-kind model that excels at PBR texture generation 🚀
🥇 SOTA commercial offering of PBR material creation, bridging the gap between AI image generation and traditional CGI
🌱 Homegrown by us at fal
🤑 Estimate your own material maps at $0.01 / map / megapixel
🖼️ 1K up to 8K material generation starting at just $0.08 for a complete five-map-plus-render material from text (and optional image)
🔍 Identify and extract materials from images with plain language
We’ve upgraded our specialized reasoning mode Gemini 3 Deep Think to help solve modern science, research, and engineering challenges – pushing the frontier of intelligence. 🧠
Watch how the Wang Lab at Duke University is using it to design new semiconductor materials. 🧵