Towards World Models in Biomedical Research
cs.AI
Submitted: 2026-06-04
Updated: 2026-08-26
License: http://creativecommons.org/licenses/by/4.0/
The gist: A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbations, disease progression and therapeutic
Terminology
Abstract
A central goal of biomedicine is to understand, predict and ultimately control the dynamic mechanisms by which biological systems respond to perturbations, disease progression and therapeutic intervention. Although foundation models and large language models have accelerated biomedical data interpretation, most current systems remain focused on static pattern recognition rather than prospective simulation of biological futures. Here we propose biomedical world models as a paradigm for AI-driven discovery. These models learn latent representations of molecular, cellular, tissue and clinical states, together with intervention-conditioned dynamics that allow future trajectories to be simulated before actions are taken. We discuss how biomedical world models could function as data engines, environment simulators and scientific planning substrates across applications including virtual cells, organoids, virtual patients and surgical simulation. We outline the data infrastructure, evaluation benchmarks, safety constraints and governance frameworks required. Biomedical world models may provide a foundation for simulation-guided, closed-loop and experimentally actionable biomedical discovery.
Sources
- GPT-4 Technical Report
- V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
- How Far Are Surgeons from Surgical World Models? A Pilot Study on Zero-shot Surgical Video Generation with Expert Assessment
- CLARITY: Medical World Model for Guiding Treatment Decisions by Simulating Context-Aware Disease Trajectories
- Is Your LLM Secretly a World Model of the Internet? Model-Based Planning for Web Agents
- Computer-Using World Model
- World Models
- Training Agents Inside of Scalable World Models
- Cosmos-H-Surgical: Learning Surgical Robot Policies from Videos via World Modeling
- Distilling the Knowledge in a Neural Network
- GAIA-1: A Generative World Model for Autonomous Driving
- Auto-Encoding Variational Bayes
- Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models
- DeepONet: Learning nonlinear operators for identifying differential equations based on the universal approximation theorem of operators
- LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
- Transformers are Sample-Efficient World Models
- Movie Gen: A Cast of Media Foundation Models
- DINOv3
- VideoGPT: Video Generation using VQ-VAE and Transformers
- ReAct: Synergizing Reasoning and Acting in Language Models
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection