SCB: SpeechConversationBench for Evaluating Multi-Turn Reasoning in Speech-to-Speech Models
cs.CL, cs.AI, cs.SD
Submitted: 2026-09-30
Updated: 2026-09-30
Terminology
Sources
- VoiceBench: Benchmarking LLM-Based Voice Assistants
- Training Verifiers to Solve Math Word Problems
- Moshi: a speech-text foundation model for real-time dialogue
- MTalk-Bench: Evaluating Speech-to-Speech Models in Multi-Turn Dialogues via Arena-style and Rubrics Protocols
- Audio MultiChallenge: A Multi-Turn Evaluation of Spoken Dialogue Systems on Natural Human Interaction
- Dynamic-SUPERB: Towards A Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech
- LLMs Get Lost In Multi-Turn Conversation
- From Awareness to Adherence: Bridging the Context Gap in Spoken Dialogue Systems via Context-Aware Decoding
- Evaluating Very Long-Term Conversational Memory of LLM Agents
- MemGPT: Towards LLMs as Operating Systems
- Thai Semantic End-of-Turn Detection for Real-Time Voice Agents
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering