ClinLens: Towards Long-Horizon LLM Agents for Longitudinal Multimodal Clinical Data Science
cs.AI
Submitted: 2026-07-28
Updated: 2026-09-26
Terminology
Sources
- Teaching Large Language Models to Self-Debug
- ScienceAgentBench: Toward Rigorous Assessment of Language Agents for Data-Driven Scientific Discovery
- BLADE: Benchmarking Language Model Agents for Data-Driven Science
- HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents
- DiscoveryBench: Towards Data-Driven Discovery with Large Language Models
- BixBench: a Comprehensive Benchmark for LLM-based Agents in Computational Biology
- PaperBench: Evaluating AI's Ability to Replicate AI Research
- Towards Generalist Biomedical AI
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection