Back to News
Argmax Inc.•April 22, 2025
TTSKit
Open Source
Audio
TTS
### TL;DR
WhisperKit is expanding into text-to-speech! TTSKit adds a new library for on-device text-to-speech using Core ML-accelerated Qwen3-TTS models (CustomVoice 0.6B and 1.7B in this first release) with real-time streaming playback on Apple Silicon. In this first PR, we're introducing the library into the WhisperKit package (WhisperKit will be renamed to reflect the new multi-Kit nature of Argmax Open-source SDK) as an optional import to add real-time TTS capabilities with a state-of-the-art open-source model, either on its own or as a complement to WhisperKit speech-to-text.
Key Insights & Metrics
Pricing
Free
Cost structure
Version
0.4
Current release version
Hardware
A12, A13, S9, S10, A16, A17 Pro, A18, A19 chips; compatible with iPhone, iPad, and Apple Watch models as specified in the config.json file.
Compute requirements
Category
Open Source
Licensing model
Region
United States
Primary region
Key Features
- On-device speech recognition
- Optimized for Apple Silicon devices
- Supports multiple Whisper models
- Efficient and accurate processing
- Seamless integration with Apple's ecosystem
Discussion
0
Upvotes
0
Downvotes
0 reviews
Sign in to leave a review
Reviews
No reviews yet. Be the first to review!