Inquesto Score: A reliability Protocol For Voice Agents
cs.CL, cs.AI
Submitted: 2026-09-24
Updated: 2026-09-24
Code: https://github.com/massabaali7/inquesto-score
Terminology
Sources
- Full-Duplex-Bench: A Benchmark to Evaluate Full-duplex Spoken Dialogue Models on Turn-taking Capabilities
- $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
- EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents
- VAmoS Bench: Voice Agent Simulation Bench
- MTVA-Bench: Evaluating the Language Model Inside Cascaded Voice Agents
- VoiceAgentEval: A Dual-Dimensional Benchmark for Expert-Level Intelligent Voice-Agent Evaluation of Xbench's Professional-Aligned Series
- Gemma 2: Improving Open Language Models at a Practical Size
- Qwen2.5 Technical Report
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering