Hunyuan3D WorldClaw
Tencent Hunyuan3D WorldClaw: text-to-3D editable game worlds (paper)
Tencent's Hunyuan team published WorldClaw, a text-to-3D system for generating editable game worlds. It is paper-only for now, with no released weights.
Interactive world models, 3D generation, and AI for games and simulations. — 20 releases covered on the show.
Tencent Hunyuan3D WorldClaw: text-to-3D editable game worlds (paper)
Tencent's Hunyuan team published WorldClaw, a text-to-3D system for generating editable game worlds. It is paper-only for now, with no released weights.
Decart's Anywear: real-time virtual try-on from any shopping site at 40ms a frame
A free Chrome extension: drag a garment from any shopping site onto your webcam feed and Decart's world model regenerates you wearing it, frame by frame at 40ms latency, with no retailer integration. Kfir Aberman demoed it live on ThursdAI, where Alex swapped his real jacket for a digital one on camera and bought a Dolce & Gabbana suit mid-interview, wearing it before it shipped. Aberman's frame: agentic commerce needs world models to close the loop between browsing and trying.
NVIDIA Lyra 2.0: single image to explorable 3D worlds, Apache 2.0
NVIDIA released Lyra 2.0 under Apache 2.0, generating persistent, explorable 3D worlds from a single image. Together with Baidu ERNIE-Image and Tencent HYWorld 2.0, it rounds out a week of open releases in the 3D-world-from-single-image race.
Tencent HYWorld 2.0 turns a single image into editable 3D scenes
Tencent released HYWorld 2.0, which converts a single image into editable 3D Gaussian Splats and meshes that are ready for Unity, Unreal, and Isaac Sim. It is one of three single-image-to-3D-world releases this week, essentially an open-source equivalent of what Fei-Fei Li's World Labs is building.
NVIDIA DLSS 5 adds a generative AI filter for photo-realistic lighting
Announced at GTC, NVIDIA's DLSS 5 introduces a new generative AI filter bringing photo-realistic lighting to RTX 50-series GPUs. It applies generative models to real-time game rendering, extending DLSS beyond upscaling and frame generation.
LingBot-World: open-source world model challenges Google Genie 3
Ant Group released LingBot-World, an open-source world model that generates 10-minute playable environments at 16fps. It positions open weights as a direct challenger to Google's closed Genie 3 in interactive world generation.
Google DeepMind launches Project Genie 3, real-time 24fps world model
Google DeepMind's Genie 3 generates interactive, controllable 3D worlds in real time at 24 frames per second, demoed live on the show with a spaceship exploration and paint persistence on walls. It ships alongside SIMA 2, a self-improving game-playing agent built on Genie 3, and is available to Gemini Ultra subscribers in the US with a one-minute session limit.
Overworld's Waypoint-1: real-time AI world model at 60fps on consumer GPUs
Overworld released Waypoint-1, a real-time AI world model that runs at 60fps on consumer GPUs. It generates interactive environments live, bringing world-model tech out of research demos and onto hardware people actually own.
SAM 3D turns single photos into 3D objects and human bodies
Released alongside SAM 3, SAM 3D reconstructs 3D objects and full human bodies from a single image with surprisingly high quality. It extends the Segment Anything family from 2D segmentation into single-image 3D reconstruction.
Odyssey V2: real-time interactive AI video you can steer as it generates
Odyssey ML launched V2 of its real-time interactive AI video experience, where the video stream is generated live and responds to user input. The panel grouped it with the week's evidence that video is becoming an interactive product surface rather than a render-and-wait demo.
World Labs RTFM renders 3D worlds in real time on a single H100
World Labs released RTFM (Real-Time Frame Model), a generative world model that renders explorable, persistent 3D worlds at interactive frame rates on a single H100 GPU. A live demo lets anyone walk through generated worlds in the browser.
Tencent launches Hunyuan 3D 3.0 with a hosted 3D studio
Tencent released Hunyuan 3D 3.0, the next version of its 3D asset generation model, available to try through a hosted 3D studio. It continues Tencent's rapid cadence of generative 3D releases.
Fei-Fei Li's World Labs presents Marble, a generative world model
World Labs, Fei-Fei Li's spatial intelligence startup, presented Marble, a generative world model that creates explorable 3D environments. The demo is treated on the show as evidence that world models are getting meaningfully closer to usable products.
Mirage debuts as the first AI-native UGC game engine
Dynamics Lab unveiled Mirage, billed as the world's first AI-native user-generated-content game engine, with real-time photorealistic playable demos powered by world-model-style generation. Alex reacted to it live as the most visibly fun demo of the week and a preview of where interactive media is headed.
Odyssey debuts real-time interactive AI video at 30 FPS
Odyssey launched interactive video: real-time AI world exploration rendered at 30 FPS, letting you walk through generated worlds as they are created. A glimpse at world-model-driven media where the video responds to you instead of just playing back.
StepFun's Step1X-3D: open two-stage framework for textured 3D assets
StepFun released Step1X-3D, an open two-stage framework for high-fidelity, controllable generation of textured 3D assets: it first synthesizes watertight geometry, then generates view-consistent textures. Trained on 2M curated meshes, the release also includes a curated dataset of 800K assets and a Hugging Face demo.
Tencent's Hunyuan 3D 2.5 jumps to 10B params with PBR textures and rigging
Tencent updated its 3D generation model to Hunyuan 3D 2.5, now boasting 10 billion parameters, up from 1B. They highlight massive leaps in precision with 1024-resolution geometry, high-quality textures with PBR support, and improved skeletal rigging for animation.
Tencent updates Hunyuan3D 2.0 with MultiView and Turbo variants
Tencent updated its Hunyuan3D 2.0 image-to-3D model with an MV (MultiView) version that conditions on multiple input views, plus a faster Turbo variant. The show highlighted it as new SOTA for 3D generation, available to try in a Hugging Face space.
Microsoft MUSE generates playable game worlds from a single second of video
Microsoft's MUSE can generate minutes of playable gameplay from just a single second of video frames and controller actions, preserving screen elements like health bars and percentages. It is based on the World and Human Action Model (WHAM) architecture, trained on a billion gameplay images from Xbox, with the model released on Hugging Face.
Tencent Hunyuan3D 2.0: SOTA open source 3D generation
Tencent released Hunyuan3D 2.0, a state-of-the-art open source 3D asset generation model on Hugging Face. It produces high-quality 3D shapes and textures and pushes open weights forward in the 3D generation category.
Follow World Models, 3D & Gaming and everything else in AI — live every Thursday.