← Back to the research library

Resources for world modeling.

129 open resources, workshops, research organizations, and technical reports from the curated list.

129 resources

Workshops & Challenges

Workshop on 4D World Models: Bridging Generation and Reconstruction @ CVPR 2026

Workshop on 4D World Models: Bridging Generation and Reconstruction @ CVPR 2026 —

ProjectSource ↗

2nd Workshop on World Models @ ICLR 2026

2nd Workshop on World Models @ ICLR 2026 —

ProjectSource ↗

Workshop on World Modeling @ Mila 2026

Workshop on World Modeling @ Mila 2026 —

ProjectSource ↗

WorldModelBench @ CVPR 2025

1st Workshop on Benchmarking World Models.

ProjectSource ↗

OpenDriveLab World Model Track @ CVPR 2025

OpenDriveLab World Model Track @ CVPR 2025 —

ProjectSource ↗

OpenDriveLab Predictive World Model Track @ CVPR 2024

OpenDriveLab Predictive World Model Track @ CVPR 2024 —

ProjectSource ↗

Argoverse 3D Occupancy Forecasting @ CVPR 2023

Argoverse 3D Occupancy Forecasting @ CVPR 2023 —

ProjectSource ↗

World Models in Physical AI @ NeurIPS 2026

Sydney; latent dynamics, generative simulation, evaluation, planning/control; co-located AV Causal Reasoning Retrieval Challenge.

ProjectSource ↗

Robot Learning with World Models: Capabilities, Frontiers, and Challenges @ NeurIPS 2026

world models and WAMs for robot reasoning, learning, and evaluation.

ProjectSource ↗

Continual World Models @ NeurIPS 2026

Sydney; world models that keep learning after deployment from observation, memory, feedback, and interaction.

ProjectSource ↗

World Models for High-Stakes Health (WMHS) @ NeurIPS 2026

Atlanta; patient world models, intervention-aware reasoning, and clinical trial simulation as a falsifiable world-model testbed.

ProjectSource ↗

Community Resources & Open Repositories

Curated Lists & Awesome Repos

Awesome World Models

Broad cross-domain curation

Awesome World Models (knightnemo)

GitHubSource ↗

Awesome World Models

General video generation, embodied AI, AD

Awesome World Models (leofan90)

GitHubSource ↗

Awesome World Model for Autonomous Driving

Driving-specific papers, benchmarks, challenges

GitHubSource ↗

Awesome World Models for Robotics

Robotics, embodied AI, VLA-adjacent work

GitHubSource ↗

Awesome-From-Video-Generation-to-World-Model

Curated trajectory from video gen to world modeling

GitHubSource ↗

Awesome Video World Models with AR Diffusion

Autoregressive diffusion recipes for scalable, consistent, interactive video world models

GitHubSource ↗

Awesome Interactive World Model

Interactive video world modeling papers, benchmarks, datasets, and resources

GitHubSource ↗

Awesome-Physical-AI

Physical AI: VLA models, world models, embodied robotic foundations

GitHubSource ↗

Awesome-WAM

World Action Models: survey, taxonomy, papers, data, and evaluation resources

ProjectSource ↗

World Model Survey Repo (Tsinghua FIB)

Survey companion: understanding world or predicting future?

GitHubSource ↗

Awesome Physics Cognition-based Video Generation

Physics plausibility in video world models

GitHubSource ↗

Awesome Robust Driving World Models

Robustness-focused driving world models

GitHubSource ↗

Awesome World Models: A Hitchhiker's Guide

Companion repo for From Masks to Worlds; emphasizes evolutionary roadmaps and memory-augmented world models

GitHubSource ↗

Learning to Model the World

Survey-centered repo for a broad AI view of world models

GitHubSource ↗

Embodied AI Paper List (HCPLab-SYSU)

Comprehensive embodied AI + world model papers

GitHubSource ↗

Embodied World Models Survey (NJU3DV)

Physical simulation + world models for embodied AI

GitHubSource ↗

Open Toolkits & Platforms

AutoLab

Provides remote, validated physical experiments and recorded action-state feedback for wireless-network world-model development; an experimentation platform rather than a completed autonomous world model.

arXivSource ↗

VidaForge

Inspectable video-data recipes with provenance and controlled pretraining studies on Wan 2.1 and V-JEPA 2.1; adjacent infrastructure for predictive pretraining.

arXivSource ↗

UnifoLM-WMA-0

Unitree framework that predicts future robot interactions for visual simulation and policy enhancement across embodiments.

ProjectGitHubSource ↗

NVIDIA Cosmos

World foundation model platform for Physical AI (robots + AD); open-weight under permissive license

GitHubarXivProjectSource ↗

NVIDIA FlashDreams

High-performance inference and serving library for interactive autoregressive video and world models

ProjectGitHubSource ↗

NVIDIA Cosmos-Predict2.5

Next-gen Cosmos WFM: flow-based, unifies Text/Image/Video2World; open checkpoints

arXivGitHubSource ↗

stable-worldmodel

Reproducible world-model research platform with data layer, baselines, planners, and OOD tasks

arXivGitHubSource ↗

minWM

Full-stack framework for building real-time interactive video world models from open video backbones

arXivGitHubSource ↗

Nano World Models

Minimalist future-video-prediction codebase with configs, eval scripts, and checkpoints

arXivGitHubSource ↗

HY-World 2.0

Open 3D world generation / simulation stack (Tencent)

GitHubSource ↗

HunyuanWorld 1.0

Text/image-to-3D explorable world generation (Tencent)

GitHubSource ↗

V-JEPA 2

Meta's latest JEPA world model for video understanding and robotic planning

GitHubSource ↗

Micro-World

AMD open-source interactive world model for game-like environments

ProjectSource ↗

LeJEPA

Lean, provable JEPA self-supervised training framework (SIGReg); ~50-line core

GitHubarXivSource ↗

Leaderboards & Benchmark Hubs

WorldRoam-Bench Leaderboard

Long-horizon stability for interactive world models (action, vision, physics, memory). Paper: WorldOdysseyBench in Benchmarks.

BenchmarksProjectSource ↗

ViewBench

Multi-view spatial-consistency benchmark for world models

arXivSource ↗

WebWorld Model Collection (Qwen)

Open WebWorld-8B/14B/32B web world-model checkpoints for agent training and lookahead search

HuggingFaceGitHubSource ↗

Datasets & Data Collections

EchoWM Unreal Data Pipeline

Unreal Engine pipeline separates physics trajectory collection from offline rendering to produce action-aligned multi-view video, with distributed scene screening and recovery; reports 8,767 hours across 1080p and 720p outputs.

arXivSource ↗

Game2World / GameCleaner

Gameplay UI taxonomy, paired synthetic videos, and in-the-wild evaluation clips support temporally consistent HUD removal; a controlled pilot evaluates cleaned footage for world-model training.

arXivGitHubSource ↗

AudioWorldSim

SoundSpaces-based open platform generates continuous binaural audio along simulated navigation rollouts, supplying reproducible acoustic training data for audio-based world models.

arXivGitHubSource ↗

EgoCS-400K

Replay-grounded egocentric Counter-Strike trajectories with video, actions, states, events, and language

arXivSource ↗

PhysEditWorld

UE5 replay dataset for physics-editable world models with gravity interventions, actions, states, and multimodal rollouts

arXivProjectSource ↗

RealWM / RealWM120K

Real-world interactive world-model data used by MagicWorld-style exploration

arXivSource ↗

WorldSimProbe

Public RoboTwin/LIBERO evaluation packages for action-conditioned world-model faithfulness probing

HuggingFaceGitHubSource ↗

Labs, Companies & Open Stacks

Who is building what, restricted to public, primary-source material already linked in this list. "Representative entries" point to sections where the full citations and badges live. Claims about unpublished internal systems are deliberately excluded.

NVIDIA (Cosmos)

World foundation model platform for Physical AI: video WFMs, driving data engines, real-time closed-loop simulation, omnimodal Cosmos 3

Representative entries in this list
Cosmos, Cosmos-Predict2.5, Cosmos 3 (§1.6, Toolkits); Cosmos-Drive-Dreams, NVIDIA OmniDreams (§1.2.1); DreamGen, FLARE (§1.3.1); SANA-WM, FlashDreams, Causal-rCM (§1.1.3, §1.6, Toolkits)
Openness
Open weights + code for the Cosmos family
§1.6Toolkits§1.2.1§1.3.1§1.1.3Source ↗

Google DeepMind (Genie)

Foundation world models for playable environments; generalist 3D agents

Representative entries in this list
Genie, Genie 2, Genie 3 (§1.1.2); SIMA (§1.3.2); MuZero (§3.1); GraphCast (§1.5.2); Physics-IQ Verified (Benchmarks)
Openness
Papers + lab reports; Genie 2/3 are blog-documented, not open-weight
§1.1.2§1.3.2§3.1§1.5.2BenchmarksSource ↗

Meta AI (JEPA)

Non-generative predictive representation program; video JEPA world models with robot planning

Representative entries in this list
I-JEPA, V-JEPA, V-JEPA 2, V-JEPA 2.1 (§2.2); LeCun's position paper (§0.1); CWM code world model (§3.6)
Openness
Open code + checkpoints (HuggingFace collection linked in Leaderboards)
§2.2§0.1§3.6LeaderboardsSource ↗

Wayve (GAIA)

Generative driving world models with fine-grained controllability

Representative entries in this list
GAIA-1, GAIA-2 (§1.2.1); lab write-ups in Blogs
Openness
Papers + technical blogs; models not released
§1.2.1BlogsSource ↗

World Labs

Spatially grounded multimodal 3D world generation

Representative entries in this list
Marble (§1.4.1); RTFM real-time frame model (Blogs)
Openness
Product + blog reports; limited technical disclosure
§1.4.1BlogsSource ↗

1X Technologies

Humanoid robotics; world models for real-robot video prediction and policy evaluation

Representative entries in this list
1x World Model Challenge (Benchmarks); OpenDriveLab WM Track co-listing
Openness
Public challenge + data; no full model paper in this list
BenchmarksSource ↗

AgiBot

Robotic manipulation world platforms, embodied data engines, embodied evaluation

Representative entries in this list
EnerVerse, EnerVerse-AC, Genie Envisioner, AgiBot World Colosseo (§1.3.1); EWMBench (Benchmarks)
Openness
Open code, datasets, and benchmarks
§1.3.1BenchmarksSource ↗

Skywork AI (Matrix)

Open real-time interactive game/world stacks; 3D world generation

Representative entries in this list
Matrix-Game 1.0/2.0 (§1.1.1); Matrix-Game 3.0 (§1.1.3); Matrix-3D (§1.4.1)
Openness
Open code + weights
§1.1.1§1.4.1Source ↗

Tencent (Hunyuan)

Explorable, mesh-based, and simulatable 3D world generation

Representative entries in this list
HunyuanWorld 1.0, HY-World 2.0 (§1.4.1, Toolkits)
Openness
Open code + weights
§1.4.1ToolkitsSource ↗

OpenAI (Sora)

Video generation framed as world simulation — the claim that started the 2024 debate

Representative entries in this list
Video generation models as world simulators (Blogs); the debate itself: Is Sora a World Simulator?, PhyWorld (Surveys, §1.5.1)
Openness
Blog post; no open model; treated here as a position, not a system entry
BlogsSurveys§1.5.1Source ↗

Tesla

Occupancy networks for planning, presented publicly in engineering talks (CVPR 2022 WAD keynote; no citable primary paper)

Representative entries in this list
Conceptual lineage is covered by the occupancy sections: §1.2.2, §2.3
Openness
Public talks only; nothing citable to list-standard
§1.2.2§2.3Source ↗

Xiaomi

Representative entries: Xiaomi EV World Model, Xiaomi-Robotics-U0, MiLA, DGGT.

Source ↗

Microsoft

Representative entries: MineWorld, Latent Spatial Memory.

Source ↗

Alibaba/Qwen and DAMO

Representative entries: WorldVLA, Qwen-RobotWorld, Qwen-AgentWorld, WorldOlympiad.

Source ↗

SenseTime

Representative entries: OpenDWM, MaskGWM, UniMLVG.

Source ↗

GigaAI

Representative entries: ReconDreamer, GigaWorld, GigaBrain.

Source ↗

Ant/Robbyant

Representative entries: LingBot-VA, LingBot-World.

Source ↗

Selected Technical Blogs & Reports

This section is intentionally selective. It favors official research-lab posts, technical reports, and a small number of high-signal explainers over general-audience trend pieces.

English — Official Labs & Primary Sources

Introducing GWM Worlds 2

Research preview of action- and camera-controlled audio-video worlds; persistent context and timestamped interactions.

Author / Source
Runway
Year
2026-09-03
BlogSource ↗

Atlas: A World Model for Spatial Intelligence

Spatial generation, 3D outputs, and space-time simulation; select-partner early access.

Author / Source
World Labs
Year
2026-09-01
BlogSource ↗

World Models

(original explainer site)

World Models (original explainer site)

Author / Source
David Ha & Jürgen Schmidhuber
Year
2018
BlogSource ↗

A Path Towards Autonomous Machine Intelligence

Author / Source
Yann LeCun, OpenReview
Year
2022
PaperSource ↗

GAIA-2: Pushing the Boundaries of Generative World Models

Author / Source
Wayve Blog
Year
2025
BlogSource ↗

Video generation models as world simulators (Sora)

Author / Source
OpenAI
Year
2024
BlogSource ↗

Genie 2: A large-scale foundation world model

Author / Source
Google DeepMind
Year
2024
BlogSource ↗

Genie 3: A new frontier for world models

Author / Source
Google DeepMind
Year
2025
BlogSource ↗

Cosmos World Foundation Models

Author / Source
NVIDIA Developer Blog
Year
2025
BlogSource ↗

SIMA: A generalist AI agent for 3D virtual environments

Author / Source
Google DeepMind
Year
2024
BlogSource ↗

The Path to Real-Time Worlds and Why It Matters

Author / Source
Over.world
Year
2025
BlogSource ↗

RTFM: A Real-Time Frame Model

Author / Source
World Labs
Year
2025
BlogSource ↗

Deep Dive into Yann LeCun's JEPA

Author / Source
Rohit Bandaru
Year
2024
BlogSource ↗

A Functional Taxonomy of World Models

Author / Source
World Labs (Fei-Fei Li et al.)
Year
2026
BlogSource ↗

Building Worlds That Train Robots (R2S2R)

Author / Source
World Labs
Year
2026
BlogSource ↗

Into the Omniverse: How Open World Models Push the Frontier of Physical AI

Author / Source
NVIDIA Blog
Year
2026
BlogSource ↗

State of World Models 2026: Taxonomy, Benchmarks and Open Challenges

Author / Source
world-models.io (Zenodo)
Year
2026
ReportSource ↗

Chinese — 官方解读 / 学术向文章

理解世界还是预测未来?清华大学世界模型全面综述

Author / Source
清华 FIB Lab / 知乎
Year
2025
BlogSource ↗

具身智能领域最新世界模型综述:250篇paper梳理主流框架

Author / Source
具身智能之心 / 知乎
Year
2025
BlogSource ↗

从专用模型到通用模型:2025年的最后一篇世界模型综述

Author / Source
知乎
Year
2025
BlogSource ↗

在2025年年初聊一下世界模型(上)

Author / Source
知乎
Year
2025
BlogSource ↗

在2025年年初聊一下世界模型(下)

Author / Source
知乎
Year
2025
BlogSource ↗

ACM综述:理解世界还是预测未来?(清华FIB Lab官方解读)

Author / Source
清华FIB Lab 官网
Year
2025
BlogSource ↗