Reddit r/LocalLLaMAAugust 26, 2026
Ran our Apache 2.0 Gepard TTS through Coval's public benchmark. 68.7 ms to first audio on one RTX 4090.
Excerpt
Coval has a public TTS leaderboard with 25 APIs on it. The harness is open source, so we just ran it against our API at https://www.nineninesix.ai and what came out. Gepard 1.0: 68.7 ms to first audio (p50) 85.8 ms at p90, worst sample was 104.8 ms 20.5 ms of silence before it starts talking 5.23% WER TTFA is lower than every number published on their board. We ran it ourselves from California, against our own endpoint, and the round trip is normalised out. Coval's own rows are not normalised, s