QuantWorlds

Models

18 models across the world-model ecosystem.

Official Agora-1 announcement artwork — "A multi-agent world model."

Agora-1

Odyssey

Multi-agent world model for multiple human or AI participants in one real-time simulation.

Public accessReal-timeInteractive
Official AlayaWorld paper teaser figure — a grid of generated interactive world scenes.

AlayaWorld

Alaya Lab

Interactive autoregressive world model with real-time camera control, prompt switching, and long-horizon memory consistency. Open-source, sustains interactive play past the one-minute mark.

Public accessOpen sourceReal-time
Official NVIDIA Cosmos 3 architecture diagram from the NVIDIA/cosmos repository.

Cosmos 3

NVIDIA

Omnimodal world-model family for physical AI spanning language, image, video, audio and action.

Public accessOpen sourceRobotics-ready
Official Google DeepMind Genie 3 announcement image — "A new era for interactive world generation."

Genie 3

Google DeepMind

General-purpose world model generating photorealistic environments from text with real-time exploration.

Public accessReal-timeInteractive
Official Tencent HY-World 1.5 (WorldPlay) project teaser.

HY-World 1.5 (WorldPlay)

Tencent Hunyuan

Tencent Hunyuan's streaming video diffusion model for real-time, interactive world modeling with long-term geometric consistency, released as WorldPlay. Given a single image or text prompt, it generates a next-chunk (16-frame) video prediction conditioned on live user keyboard/mouse actions, dynamically reconstituting context memory from past frames to keep the world geometrically consistent over long sessions. Runs at up to 24fps in first- or third-person, across real-world and stylized scenes. Open-sourced (code, weights, and full training pipeline). Predecessor to, and released separately from, the later HY-World 2.0.

Public accessOpen sourceReal-time

1 world on QuantWorlds

Official Tencent HY-World 2.0 project teaser showing generated 3D avatars and environments.

HY-World 2.0

Tencent Hunyuan

Multi-modal world model framework for world generation and world reconstruction — accepts text, single-view images, multi-view images, and video to produce editable, persistent 3D world representations (meshes / 3D Gaussian Splatting) compatible with game engines (Blender, Unity, Unreal).

Public accessOpen sourcePersistent worlds
Official Hugging Face social thumbnail for Lightricks/LTX-2.5.

LTX-2.5

Lightricks

Lightricks' 22-billion-parameter open-weights audio-video generation model, released August 11, 2026. Turns text, image, and video inputs into synchronized, high-fidelity video and audio in a single pass, including native multishot generation (connected scenes with consistent character/style across cuts). Positioned by Lightricks and press coverage as a "world model" for video, robotics and simulation, distinct from this catalogue's live-interactive world models (Marble, MIRA, PAN, WorldPlay) — see research notes for the classification nuance.

Public accessOpen sourceAPI available

Lucy 2.5

Decart

Real-time world editing/transformation model for live immersive video experiences.

Public accessReal-timeInteractive

Marble 1.0 Draft

World Labs

Fast Marble variant for rapid exploration.

Public accessInteractiveAPI available

Marble 1.1

World Labs

World Labs world-generation model with improved quality at fixed generation cost.

Public accessInteractiveAPI available

1 world on QuantWorlds

Marble 1.1 Plus

World Labs

Advanced Marble model from World Labs for larger persistent 3D worlds.

Public accessInteractiveAPI available
Official qualitative result frame from the Marionette paper — a game world model interacting with a creature in a cave environment.

Marionette

Alaya Lab

Marionette predicts an explicit, interpretable 3D world state, renders its geometry with a graphics operator that has no learnable parameters, and asks a video-diffusion model for one thing only: appearance. A world model for interactive games with articulated characters, trained on recordings from a commercial action game (drawn from the WildWorld corpus).

Official Matrix-Game 3.0 architecture diagram from Skywork AI.

Matrix-Game 3.0

Skywork AI

Real-time and streaming interactive world model with long-horizon memory, for controllable game world generation.

Open sourceReal-timeInteractive
Official MIRA architecture diagram showing the representation codec encoder/decoder pipeline.

MIRA

General Intuition

Multiplayer Interactive World Models with Representation Autoencoders — a 5-billion-parameter diffusion transformer (with a 600M-parameter video representation codec) that simulates 2v2 Rocket League matches in real time at 20 FPS, taking action streams from up to 4 agents at once. Trained on ~10,000 hours of bot-generated 2v2 matches, without an explicit physics engine or 3D representation.

Real-timeInteractiveMulti-agent

1 world on QuantWorlds

Official Oasis 3 branding image from Decart, arranging driving-scene photos in a circular composition.

Oasis 3

Decart

Interactive world model for physical AI with controllable multi-view simulation in real time.

Public accessReal-timeInteractive
Official Odyssey-2 announcement image from Odyssey.

Odyssey-2

Odyssey

General-purpose real-time world model generating interactive simulations from text or image prompts.

Public accessReal-timeInteractive
Official PAN project hero image from MBZUAI — a translucent globe over a grass and water landscape.

PAN

MBZUAI

General, interactable, long-horizon world simulation model — can be manipulated at intermediate steps and maintains consistency over long time horizons, evaluated as competitive with leading commercial world models.

Interactive

1 world on QuantWorlds

Starchild-1

Odyssey

Odyssey world model exploring richer multimodal interaction beyond visual observation alone.