TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context
TontaubeV1: Streaming Text-to-Speech with Hierarchical Codec Modeling and Bounded Context. It centres on Inference, and also names Benchmarks and ElevenLabs. Reported by arXiv. Bharat Hunt files it under AI Hardware — the section covering chips, accelerators, data centres, on-device inference and supply.