Kokoro TTS is a lightweight AI text-to-speech model with 82 million parameters, built on the StyleTTS 2 architecture. It delivers high-quality, natural-sounding voice synthesis while being resource-efficient. The model supports multiple languages including American English, British English, French, Korean, Japanese, and Mandarin, with customizable voicepacks for different tones and styles.
Key Features
- 82M Parameter Efficiency: Maintains high-quality speech with minimal computational resources.
- Multilingual Support: Generates speech in six languages for global projects.
- Customizable Voicepacks: Offers lifelike voice options that can be tailored to project needs.
- Automatic Content Segmentation: Detects chapters and sections for organized audio output.