Data Scale, Not Latency, Shapes Cross-Lingual Encoder Transfer in Streaming ASR
cs.AI
Submitted: 2026-06-23
Updated: 2026-09-13
Comments: Accepted at IEEE SLT 2026
Code: https://github.com/huggingface/open
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Towards scalable efficient on-device ASR with transfer learning
- Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
- Sequence Transduction with Recurrent Neural Networks
- Scaling Laws for Transfer
- NeMo: a toolkit for building AI applications using Neural Modules
- Pushing the Limits of On-Device Streaming ASR: A Compact, High-Accuracy English Model for Low-Latency Inference
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection