HappyWorld-Bench
cs.CV
Submitted: 2026-09-21
Updated: 2026-09-21
Code: https://github.com/ZiYang-xie/WorldGen
Project page: https://oasis-model.github.io
Terminology
Sources
- World Models
- Mastering Diverse Domains through World Models
- Matrix-Game 3.0: Real-Time and Streaming Interactive World Model with Long-Horizon Memory
- SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer
- Advancing Open-source World Models
- VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
- WorldScore: A Unified Evaluation Benchmark for World Generation
- WorldMark: A Unified Benchmark Suite for Interactive Video World Models
- iWorld-Bench: A Benchmark for Interactive World Models with a Unified Action Generation Framework
- WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation
- MBench: A Comprehensive Benchmark on Memory Capability for Video World Models
- WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models
- PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives
- EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models
- WorldModelBench: Judging Video Generation Models As World Models
- MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models
- RoboTrustBench: Benchmarking the Trustworthiness of Video World Models for Robotic Manipulation
- WorldArena: A Unified Benchmark for Evaluating Perception and Functional Utility of Embodied World Models
- RoboWM-Bench: A Benchmark for Evaluating World Models in Robotic Manipulation
- Eval3D: Interpretable and Fine-grained Evaluation for 3D Generation
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models