To the line
Directions · GENERATIVE SIMULATION

World models

In one viewNot a video clip but a working world you can enter and act in: turn the camera and the model remembers what was behind you. Such worlds are where robots will train.

Environment simulation instead of passive video. The same technology is the trainer where physical skills are cheaper and safer to learn than on real hardware.

Status
TypeDirections
Marker
Events in dossier8
Development chronology

Researched

2025-08

Genie 3

About this eventReal-time interactive worlds with scene consistency lasting several minutes.

2026-01-29

Project Genie

About this eventPlayable generated worlds opened to some Google AI Ultra subscribers in the US.

2026-04-26

Sora app shut down

About this eventThe standalone app closed and the API ends on 24 Sep 2026 — video generation folds into general models.

2026-05

Gemini Omni

About this eventGoogle I/O 2026 unveiled a model that creates video from any input.

2026-05-31

NVIDIA Cosmos 3

About this eventAn open physical-AI omnimodel combined vision, world simulation and action generation in one architecture.

In progress

now

Physics by eye

About this eventObjects fall and bounce plausibly, but the model guesses physics rather than computing it. On long scenes it shows.

Planned

planned

Crossing into reality

About this eventA skill learned in simulation must work on real hardware. The sim-to-real gap is the main obstacle.

Distant horizons

ahead

The world as a workplace

About this eventA fully traversable environment training not just robots but agents — from driving to running a factory.

Sources and research

Primary material behind this dossier: papers, lab publications and official reports.

Enabled by