Back to News
Dongchao Yang•February 4, 2026
UniAudio 2.0
Open Source
Audio
TTS
### TL;DR
UniAudio 2.0 is a unified audio language model designed to handle text, speech, sound, and music. It introduces ReasoningCodec, a discrete audio codec that factorizes audio into reasoning and reconstruction tokens, enhancing both understanding and generation tasks. The model is trained on 100 billion text tokens and 60 billion audio tokens, demonstrating strong performance across various audio tasks.
Key Insights & Metrics
Pricing
Free
Cost structure
Version
2.0
Current release version
Hardware
NVIDIA A100 GPU, 64GB RAM
Compute requirements
Category
Open Source
Licensing model
Region
China
Primary region
Key Features
- ReasoningCodec for audio factorization
- Unified autoregressive architecture for text and audio
- Trained on extensive text and audio datasets
- Strong few-shot and zero-shot generalization
- Competitive performance across speech, sound, and music tasks
Discussion
0
Upvotes
0
Downvotes
0 reviews
Sign in to leave a review
Reviews
No reviews yet. Be the first to review!