Oneira: From Open-Ended Generation to Open-World Interaction in Video World Models
cs.CV
Submitted: 2026-10-01
Updated: 2026-10-01
Code: https://github.com/MiniMax-AI/MiniMax-H3
Project page: https://madaoer.github.io/projects/oneira
Terminology
Sources
- NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation
- MASS: Multiplayer World Models with Authoritative Shared State
- Geometric Context Transformer for Streaming 3D Reconstruction
- World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration
- Code World Model: Coding Agent as World Brain
- Infinite Worlds with Versatile Interactions
- Coarse-to-Real: Generative Rendering for Populated Dynamic Scenes
- MineWorld: a Real-Time and Open-Source Interactive World Model on Minecraft
- World Models
- Mastering Diverse Domains through World Models
- Matrix-game 2.0: An open-source, real-time, and streaming interactive world model
- RELIC: Interactive Video World Model with Long-Horizon Memory
- LoRA: Low-Rank Adaptation of Large Language Models
- Generative World Renderer
- DreamGen: Unlocking Generalization in Robot Learning through Video World Models
- ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
- Sekai: A Video Dataset towards World Exploration
- StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation
- WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models