DuplexSpeechBench-Document Grounding: Benchmarking Document Grounding and Hallucinations in Voice Agents
cs.CL, cs.AI
Submitted: 2026-09-29
Updated: 2026-09-29
Code: https://github.com/hexgrad/kokoro
Project page: https://dsb-dg.github.io/Abstract
Terminology
Sources
- MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction
- Moshi: a speech-text foundation model for real-time dialogue
- Spoken SQuAD: A Study of Mitigating the Impact of Speech Recognition Errors on Listening Comprehension
- Full-Duplex-Bench-v3: Benchmarking Tool Use for Full-Duplex Voice Agents Under Real-World Disfluency
- Full-Duplex-Bench: A Benchmark to Evaluate Full-duplex Spoken Dialogue Models on Turn-taking Capabilities
- FD-Bench: A Full-Duplex Benchmarking Pipeline Designed for Full Duplex Spoken Dialogue Systems
- $\tau$-Voice: Benchmarking Full-Duplex Voice Agents on Real-World Domains
- SALMONN-omni: A Codec-free LLM for Full-duplex Speech Understanding and Generation
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering