NineNineSix
2026•NineNineSix
Gepard 1.0 v1.0
Open Source
text-to-speech
voice-cloning
Gepard 1.0 is an autoregressive, prosody-aware text-to-speech model designed specifically for real-time conversational AI. It utilizes a Qwen3.5-based backbone to generate audio in single-pass frames, enabling ultra-low latency and high-throughput streaming suitable for voice agents.
Single-pass, streaming-first frame generation for low-latency dialogue
vLLM-native support allowing for high-throughput concurrent sessions
Zero-shot voice cloning integrated into the prefill stage
PricingFree
Version1.0
2025•NineNineSix
KaniTTS2-en v1.0
Open Source
Audio
TTS
KaniTTS2-en is an open-source text-to-speech model developed by NineNineSix, optimized for real-time conversational AI applications. It utilizes a two-stage pipeline combining a large language model with an efficient audio codec to deliver high-quality, low-latency speech synthesis.
Two-stage pipeline combining LLM and FSQ audio codec
Model size of 400 million parameters
Supports English language with a sample rate of 22kHz
PricingFree
Version1.0