TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context

arXiv:2609.08703 · cs.SD, cs.CL, cs.LG · Submitted 2026-09-08 · Read on arXiv

cs.SD, cs.CL, cs.LG

Submitted: 2026-09-08

Updated: 2026-09-08

Comments: 13 pages, 2 figures, 2 tables. Both authors contributed equally

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Related papers