How Well Can LLMs Simulate Real Learner Evaluations of Educational Feedback?
cs.CL
Submitted: 2026-09-28
Updated: 2026-09-28
Terminology
Sources
- Qwen3-VL Technical Report
- A Systematic Analysis of the Impact of Persona Steering on LLM Capabilities
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
- Investigating Learner-Aware Design of LLM-Generated Educational Feedback
- Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization
- Dean of LLM Tutors: A Framework for Automated Quality Review of AI-generated Feedback
- OpenAI GPT-5 System Card
- Gemma 4 Technical Report
- Automatic Generation of Question Hints for Mathematics Problems using Large Language Models in Educational Technology
- Simulating LLM-to-LLM Tutoring for Multilingual Math Feedback
- EduPersona: Benchmarking Subjective Ability Boundaries of Virtual Student Agents
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering