NineNineSix

2026NineNineSix

Gepard 1.0 v1.0

Open Source
text-to-speech
voice-cloning

Gepard 1.0 is an autoregressive, prosody-aware text-to-speech model designed specifically for real-time conversational AI. It utilizes a Qwen3.5-based backbone to generate audio in single-pass frames, enabling ultra-low latency and high-throughput streaming suitable for voice agents.

Single-pass, streaming-first frame generation for low-latency dialogue
vLLM-native support allowing for high-throughput concurrent sessions
Zero-shot voice cloning integrated into the prefill stage
PricingFree
Version1.0
2025NineNineSix

KaniTTS2-en v1.0

Open Source
Audio
TTS

KaniTTS2-en is an open-source text-to-speech model developed by NineNineSix, optimized for real-time conversational AI applications. It utilizes a two-stage pipeline combining a large language model with an efficient audio codec to deliver high-quality, low-latency speech synthesis.

Two-stage pipeline combining LLM and FSQ audio codec
Model size of 400 million parameters
Supports English language with a sample rate of 22kHz
PricingFree
Version1.0

🚀 Join the AI dev community — follow us everywhere

© 2026 MARKTECHPOST AI MEDIA INC. All rights reserved.Terms & ConditionsPrivacy Policy
Beta Mode