Kimi K3 combines strong 3D reasoning, coding, and vision capabilities to turn concepts, images, and videos into fully playable interactive experiences.
Kimi K3 achieves true "vision in the loop" by seamlessly iterating between code and live screenshots
Can video generation models do for vision what LLMs did for language?
Introducing GenCeption from @GoogleDeepMind: one feed-forward video model for various vision tasks — SOTA, data-efficient, and emerging behaviors (ECCV 2026)
🌐 https://t.co/3rwVgTolZz
(1/8)
I will also be there presenting LuxRemix!
Looking forward to your visit 🤗
BTW, I just released our massive 360 synthetic rendering dataset w/ light decomposition. Please check the thread.
Today, we’re presenting LuxRemix at #CVPR2026 — 10:45 poster session, poster #102!
Decompose indoor lighting into individual light sources from one image, then remix it (on/off, color, intensity) with real-time 3DGS relighting.
Come say hi 👋
https://t.co/snCmEm58xt
World models are moving beyond offline generation towards interactive, real-time experiences.
Introducing ⚡FlashDreams⚡: an open-source high-performance inference and serving library built for autoregressive world models:
🔥 Up to 3.10× faster LingBot-World inference
🔥 Up to 2.12× faster Self-Forcing inference
🔥 Up to 1.40× faster Wan2.1 inference
🔥 8 integrated models
🔥 Multi-GPU, streaming, low-latency serving
🔥 Agentic skills that teach you how to use it
FlashDreams is designed for a new generation of AI systems that continuously evolve over time while responding to user interactions. It powers applications across robotics, autonomous vehicle simulation, gaming, and virtual worlds.
Github: https://t.co/xM8LuPaRTS
Docs: https://t.co/IInORNIzy3
Research page: https://t.co/mZ6TLQSpIO
Join the #flashdreams Discord channel at https://t.co/GGOQ0k7liY
FlashDreams is also the runtime backbone behind NVIDIA OmniDreams (https://t.co/PLUt55gxxh)
1/n
#AI #WorldModels #FastInference #PhysicalAI #OpenSource #NVIDIA
🔥 Check out our latest work!
🚀 Introducing Gamma-World: a generative multi-agent world model that scales beyond two players and supports real-time streaming at 24 FPS.
Introducing VGGT-Ω: scaling feed-forward reconstruction across static and dynamic scenes, and studying whether the learned geometric representations transfer beyond reconstruction.
Fortunate to help @yxue_yxue on this project during our Meta internship 😃.
GeoRelight follows the research trajectory of DifffusionRenderer and UniRelight and delivers a significant quality improvement in relighting and geometry reconstruction for general human captures.
#CVPR2026 Highlight
How to make relighting more photorealistic? Make reconstruction happening together!
GeoRelight jointly resolves Geometry, Instrinsics, and Relighting, and proves they have mutual benefit (Geometry helps you know shadow and shading)
https://t.co/fpMUvXctiP
We open-sourced the code and model for UniRelight! 🎉
Given an input video and a target lighting configuration, our method jointly predicts a relit video and its corresponding albedo.
Code: https://t.co/4zF94saWvo
Model: https://t.co/d8i66UyvhU
The code is out for LuxDiT! 🎉
LuxDiT allows you to estimate high-quality HDR environment maps from images or videos. You can use our released Gradio web UI to quickly try your input.
🗒️Code: https://t.co/YYUIqGTpjs
💾Model: https://t.co/EZKKs46IVh
💡 Introducing LuxDiT: a diffusion transformer (DiT) that estimates realistic scene lighting from a single image or video.
It produces accurate HDR environment maps, addressing a long-standing challenge in computer vision.
🔗Paper: https://t.co/6cW6WlREBl
Spark 2.0 is here! 🚀
We’re redefining what’s possible on the web with a streamable LoD system for 3D Gaussian Splatting.
Built on Three.js, you can now stream massive 100M+ splat worlds to any device from mobile to VR using WebGL2. All open-source.
Dive into the tech 👇
We scaled up Lyra to generate explorable 3D worlds! 🚀
Introducing Lyra 2.0 — turning a single image into a 3D world you can walk through, look back, and even drop a robot into 🤖
Code and Model available today!
🌐 Website: https://t.co/plBxCoWkNn
(1/N)
𝗜𝗻𝘁𝗿𝗼𝗱𝘂𝗰𝗶𝗻𝗴 𝗚𝗲𝗻𝗲𝗿𝗮𝘁𝗶𝘃𝗲 𝗪𝗼𝗿𝗹𝗱 𝗥𝗲𝗻𝗱𝗲𝗿𝗲𝗿
A new toolkit, dataset, and baseline for world-scale rendering:
- collect G-buffers from AAA games
- scale data for rendering complex world scenes
- improve rendering performance
- enable game effect editing
My dear front-end developers (and anyone who’s interested in the future of interfaces):
I have crawled through depths of hell to bring you, for the foreseeable years, one of the more important foundational pieces of UI engineering (if not in implementation then certainly at least in concept):
Fast, accurate and comprehensive userland text measurement algorithm in pure TypeScript, usable for laying out entire web pages without CSS, bypassing DOM measurements and reflow
I also took a picture of this classic place for the teaser demo of our #CVPR26 paper LuxRemix (https://t.co/HCVianYn7g) 😄
I used to love to sit there on Saturday afternoon, hoping Einstein could give me some research ideas 🤣
Credits also to @c_richardt for composing the vid.
Announcing NVIDIA DLSS 5, an AI-powered breakthrough in visual fidelity for games, coming this fall.
DLSS 5 infuses pixels with photorealistic lighting and materials, bridging the gap between rendering and reality.
Learn More → https://t.co/yHON3nGyxE
📢Introducing 360Anything, our method for lifting any perspective image or video to gravity-aligned 360° panoramas without using any camera or 3D information. This enables consistent novel view synthesis and 3D scene reconstruction.
Project page: https://t.co/qTOEip0Jw2
🧵
Excited to share our new work: LuxRemix ✨
We leverage diffusion models to decompose complex light transport into individual sources. You can interactively remix room lighting in 2D images or 3D GSplats 💡🎛️
Try the interactive light controls here: 🔗 https://t.co/h98dX44frl
We’re excited to share LuxRemix: interactive light editing for indoor scenes! 🏠💡
Capture a room once, then turn individual lights on/off, change colors, and adjust intensity – all in real-time 3D from any viewpoint.
💡 https://t.co/MZNEDpEW68
📄 https://t.co/OldK7jZbJi