Multilingual Medical Reasoning for Question Answering with Large Language Models
cs.CL, cs.AI
Submitted: 2025-12-05
Updated: 2026-09-01
Code: https://github.com/vllm-project/vllm
Terminology
Sources
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- The Llama 3 Herd of Models
- Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL
- O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
- GPT-4o System Card
- OpenAI o1 System Card
- Multi-Step Reasoning with Large Language Models, a Survey
- O1 Replication Journey: A Strategic Progress Report -- Part 1
- Qwen2.5 Technical Report
- MedGemma Technical Report
- Gemma 3 Technical Report
- Disentangling Reasoning and Knowledge in Medical Large Language Models
- Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and Applications
- Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
- MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs
- Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
- Qwen3 Technical Report
- Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering