Back to News
Fish Audio•March 10, 2026
Fish Audio S2 Pro
Open Source
Audio
### TL;DR
Fish Audio S2 Pro is an advanced text-to-speech (TTS) model developed by Fish Audio, offering fine-grained inline control over prosody and emotion. Trained on over 10 million hours of audio data across more than 80 languages, it combines reinforcement learning alignment with a dual-autoregressive architecture to deliver high-quality, expressive speech synthesis.
Key Insights & Metrics
Pricing
Free for research and non-commercial use; commercial use requires a separate license from Fish Audio.
Cost structure
Version
1.0
Current release version
Hardware
NVIDIA H200 GPU recommended for optimal performance; other GPUs may also be compatible.
Compute requirements
Category
Open Source
Licensing model
Region
Unknown
Primary region
Key Features
- Fine-grained inline control over prosody and emotion
- Supports over 80 languages
- Dual-autoregressive architecture for efficient inference
- Real-time streaming performance on NVIDIA H200 GPU
- Open-source model weights and fine-tuning code
Discussion
0
Upvotes
0
Downvotes
0 reviews
Sign in to leave a review
Reviews
No reviews yet. Be the first to review!