Back to News
Kyutai•February 12, 2026
Hibiki-Zero
Open Source
Audio
TTS
AGI
### TL;DR
Hibiki-Zero is an advanced model for simultaneous speech-to-speech translation that eliminates the need for word-level aligned data, simplifying the training pipeline and enabling seamless scaling to diverse languages. It achieves state-of-the-art performance in translation accuracy, latency, voice transfer, and naturalness across multiple tasks.
Key Insights & Metrics
Pricing
Free
Cost structure
Version
1.0
Current release version
Hardware
GPU with at least 16GB VRAM recommended
Compute requirements
Category
Open Source
Licensing model
Region
France
Primary region
Key Features
- Simultaneous speech-to-speech translation without word-level aligned data
- State-of-the-art performance in translation accuracy and latency
- Preserves speaker identity and speech naturalness
- Adaptable to new languages with less than 1000 hours of speech data
- Provides model weights, inference code, and a 45-hour multilingual speech benchmark for evaluation
Discussion
0
Upvotes
0
Downvotes
0 reviews
Sign in to leave a review
Reviews
No reviews yet. Be the first to review!