Static HTML snapshot of the model record for crawlers and no-JS readers.
| Attribute | Value |
|---|---|
| Model | Genie 3 |
| Lab / Organization | DeepMind |
| Category | Generative World Model |
| Subtype | Interactive 3D World Model |
| World Model Type | Real-time interactive 3D environment generation |
| Primary Domain | 3D World Generation |
| Architecture | Autoregressive latent world model with spatiotemporal transformer |
| Modality | Text → Interactive 3D Environment |
| Training Method | Large-scale internet video and 3D data pre-training with action-conditioned generation |
| Status | active |
| Year | 2025 |
| Performance Index | 89/100 (medium confidence, v1.1) |
Main editorial body preserved directly in static HTML.
Genie 3 is a general-purpose world model from Google DeepMind that generates an unprecedented diversity of interactive environments from text prompts. Users can navigate generated worlds in real time at 24 frames per second, with consistency maintained for several minutes at 720p resolution. Building on Genie 1 and Genie 2, this third generation dramatically expands the scope of generated worlds, from natural landscapes and ecosystems to architectural spaces and fantastical environments. The model demonstrates emergent understanding of physics, lighting, object permanence, and spatial relationships. Project Genie, the public product built on Genie 3, was made available to Google AI Ultra subscribers in January 2026.
Genie 3 is a real-time interactive 3d environment generation developed by Google DeepMind in 2025 for 3d world generation.
Short extractable facts for answer engines and no-JS readers.
| Signal | Value |
|---|---|
| Definition | Genie 3 is a real-time interactive 3d environment generation developed by Google DeepMind in 2025 for 3d world generation. |
| Short Description | Google DeepMind's general-purpose world model that generates interactive 3D environments from text prompts in real time at 24fps. |
| Benchmark Rows | 2 |
| FAQ Entries | 2 |
| Related Models | 4 |
| Related Guides | 0 |
| Related Research Topics | 3 |
| Last Updated | 2026-04-10 |
Key capabilities associated with this model.
Representative applications attached to this model record.
Balanced assessment surfaced in static HTML.
Primary references preserved in static HTML for citation extraction.
| Reference | Link |
|---|---|
| Parker-Holder & Fruchter, 2025. Genie 3: A New Frontier for World Models. Google DeepMind Blog. | Open source |
Nearby models linked from the current editorial record.
| Model | Category | World Model Type | Index v1.1 |
|---|---|---|---|
| Genie 2 | Generative World Model | Generative environment model | 79/100 |
| Genie | Generative World Model | Action-controllable generative model | 57/100 |
| NVIDIA Cosmos | Foundation World Model | Video world foundation model | 87/100 |
| OASIS | Generative World Model | Real-time playable world model | 66/100 |
Side-by-side comparisons already connected to this model.
| Comparison | Matchup | Summary |
|---|---|---|
| Genie 3 vs Genie 2 | Genie 3 vs Genie 2 | Two generations of DeepMind's interactive world model. Genie 2 generates 3D environments from single images; Genie 3 generates them from text prompts in real time at 24fps with far greater diversity and consistency. |
| Sora vs Genie 3 | Sora vs Genie 3 | Two leading generative world models with opposite design goals: Sora targets long, cinematic, non-interactive clips from text, while Genie 3 trades fidelity for real-time controllable worlds you can actually play in. |
| Pandora vs Genie 2 | Pandora vs Genie 2 | Both generate explorable 3D-feeling environments, but with different control surfaces: Pandora accepts free-form text actions through an LLM backbone, while Genie 2 conditions on a single seed image and learned latent actions. |
| Genie 3 vs NVIDIA Cosmos | Genie 3 vs NVIDIA Cosmos | Two green-index frontier systems with different ambitions. Genie 3 is a real-time text-to-world interactive generator, while NVIDIA Cosmos is a broad physical-AI platform optimized for simulation infrastructure, robotics, and industrial world modeling. |
| Genie 3 vs V-JEPA 2 | Genie 3 vs V-JEPA 2 | Two green-index leaders that represent different frontier philosophies. Genie 3 is an interactive generative world model that turns text into playable environments, while V-JEPA 2 is a self-supervised latent predictor optimized for physical reasoning and zero-shot robot planning. |
| NVIDIA Cosmos vs V-JEPA 2 | NVIDIA Cosmos vs V-JEPA 2 | Two green-index foundation-scale leaders with different views of world modeling. Cosmos emphasizes a platform for physical-AI simulation and generation, while V-JEPA 2 emphasizes self-supervised predictive representations for visual understanding and robot control. |
| PlayWorld vs V-JEPA 2 | PlayWorld vs V-JEPA 2 | Two green-index models pushing robotics-relevant world understanding in different ways. PlayWorld is a robot-play simulator for manipulation and policy improvement, while V-JEPA 2 is a self-supervised video predictor optimized for physical reasoning and zero-shot robot planning. |
| DreamerV3 vs PlayWorld | DreamerV3 vs PlayWorld | Two green-index leaders for acting under learned dynamics, but with different centers of gravity. DreamerV3 is the canonical imagination-based general RL agent, while PlayWorld is a manipulation-centric robot simulator trained from autonomous play data. |
Connected research areas surfaced directly in static HTML.
| Topic | Summary |
|---|---|
| Video World Models | How video world models learn physics, temporal consistency, and interactive simulation from large-scale video, from Sora and Genie to Cosmos and V-JEPA. |
| Diffusion World Models | How diffusion world models generate future states, preserve richer visual detail, and power video simulation systems such as DIAMOND, Sora, and Cosmos. |
| Language-Conditioned World Models | How language-conditioned world models use text prompts or natural-language actions to control simulation, planning, and embodied behavior across Pandora, 3D-VLA, RT-2, and hybrid systems. |
Recent timeline events connected to this model.
| Event | Published | Source | Summary |
|---|---|---|---|
| Google DeepMind unveils Genie 3: real-time interactive world model at 1080p/24fps | 2026-04-22 | Google DeepMind | DeepMind announces Genie 3, a real-time foundation world model generating interactive 3D environments at 720p and 24 fps from text prompts. |
FAQ answers rendered directly into static HTML for extractable responses.
Genie 3 introduces real-time generation at 24fps (vs offline in Genie 2), supports text prompts instead of just image conditioning, generates far more diverse environments, and is publicly accessible via Project Genie.
Yes: Google AI Ultra subscribers in the U.S. can access Project Genie, a research prototype built on Genie 3, to create and explore AI-generated worlds.
Short extractable summary preserved directly in static HTML.
Editorial provenance and refresh policy preserved directly in static HTML.
Published by world-models.io editorial board.
Lead editor Tyler D. - Technical editor, methodology and benchmark analysis.
This model page synthesizes primary papers, official model pages, benchmark evidence, and related world-models.io context into a reference resource.
Each editorial page is assembled from primary sources, normalized into extractable summaries, checked for factual drift, and reviewed before publication or major refreshes. Last reviewed: 2026-04-10.
Pages are refreshed when a new paper, benchmark, release, architecture update, or stronger primary source materially changes the answer a reader or AI system should retrieve.
Each page links back to relevant primary sources and keeps a stable canonical URL so readers can verify claims, trace context, and reference the most up-to-date version. See the editorial policy.
Primary model and lab sources embedded in static HTML.