A Patient Simulation Framework for Risk Assessment of Conversational Healthcare AI: Evaluation of an Antidepressant Decision Aid
cs.CL
Submitted: 2026-02-11
Updated: 2026-09-09
Code: https://github.com/malexandersalazar/xlm-roberta-base-cls-depression
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Evaluating LLM-based Agents for Multi-Turn Conversations: A Survey
- Measuring Recommender System Effects with Simulated Users
- PatientSim: A Persona-Driven Simulator for Realistic Doctor-Patient Interactions
- A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions
- SimSUM: Simulated Benchmark with Structured and Unstructured Medical Records
- Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education
- Modeling Challenging Patient Interactions: LLMs for Medical Communication Training
- MATRIX: Multi-Agent simulaTion fRamework for safe Interactions and conteXtual clinical conversational evaluation
- HAICOSYSTEM: An Ecosystem for Sandboxing Safety Risks in Human-AI Interactions
- Learning diverse attacks on large language models for robust red-teaming and safety tuning
- Automatic Interactive Evaluation for Large Language Models with State Aware Patient Simulator
- LLMaAA: Making Large Language Models as Active Annotators
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering