Miso One Open-Sources Voice Model: 8B Parameters, 110ms Latency, One-Shot Voice Cloning
Miso One 开源语音模型:8B 参数、110ms 延迟、一次语音克隆
Miso One released an 8B-parameter open-weight TTS model with one-shot voice cloning from a short sample, 110ms inference latency, GitHub self-hosting without an API, and local audio data handling; the post says API access is coming but does not disclose pricing or launch timing.
Why it matters: HKR-H/K/R all pass, but this is a single X-sourced launch with no benchmark suite, license detail, or third-party reproduction. The 8B, 110ms, self-hosted open TTS facts clear featured, not higher.