OpenTumorBoard: A Real-World Benchmark of Multidisciplinary Tumor Board Discussion Trajectories
cs.CL, cs.LG
Submitted: 2026-09-26
Updated: 2026-09-29
Terminology
Sources
- Demo: Healthcare Agent Orchestrator (HAO) for Patient Summarization in Molecular Tumor Boards
- Multimodal Clinical Benchmark for Emergency Care (MC-BEC): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine
- MedAgentBench: A Realistic Virtual EHR Environment to Benchmark Medical LLM Agents
- What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoning
- Understanding R1-Zero-Like Training: A Critical Perspective
- AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- MedAgentGym: A Scalable Agentic Training Environment for Code-Centric Reasoning in Biomedical Data Science
- MedAgentBoard: Benchmarking Multi-Agent Collaboration with Conventional Methods for Diverse Medical Tasks
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering