MeteoVerse: Unified Weather-Controllable Video World Model
cs.CV
Submitted: 2026-09-29
Updated: 2026-09-29
Project page: https://meteoverse.github.io
Terminology
Sources
- ReCamMaster: Camera-Controlled Generative Rendering from A Single Video
- Qwen3-VL Technical Report
- Genie: Generative Interactive Environments
- Uni3C: Unifying Precisely 3D-Enhanced Camera and Human Motion Controls for Video Generation
- CameraCtrl: Enabling Camera Control for Text-to-Video Generation
- GAIA-1: A Generative World Model for Autonomous Driving
- AutoAWG: Adverse Weather Generation with Adaptive Multi-Controls for Automotive Videos
- ClimateNeRF: Extreme Weather Synthesis in Neural Radiance Field
- Controllable Weather Synthesis and Removal with Video Diffusion Models
- WeatherEdit: Controllable Weather Editing with 4D Gaussian Field
- Advancing Open-source World Models
- ACDC: The Adverse Conditions Dataset with Correspondences for Robust Semantic Driving Scene Perception
- SHIFT: A Synthetic Driving Dataset for Continuous Multi-Task Domain Adaptation
- TransWeather: Transformer-based Restoration of Images Degraded by Adverse Weather Conditions
- Diffusion Models Are Real-Time Game Engines
- VGGT-$\Omega$
- MotionCtrl: A Unified and Flexible Motion Controller for Video Generation
- WeatherCity: Urban Scene Reconstruction with Controllable Multi-Weather Transformation
- Video Adverse-Weather-Component Suppression Network via Weather Messenger and Adversarial Backpropagation
- Genuine Knowledge from Practice: Diffusion Test-Time Adaptation for Video Adverse Weather Removal
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models