# StyleTTS2

> Human-level TTS via style diffusion and adversarial training

- Category: [Foundation Models](https://ailandscape.org/category/foundation-models) › Speech & Audio Models
- Homepage: https://styletts2.github.io
- Repository: https://github.com/yl4579/StyleTTS2
- Tags: tts, voice, oss
- Added to the landscape: 2026-03-18

## Similar tools in Speech & Audio Models

- [AudioLDM 2](https://ailandscape.org/tool/audioldm-2): Latent diffusion model for audio and music generation
- [Bark](https://ailandscape.org/tool/bark): Open-source text-to-audio model by Suno AI
- [Kokoro](https://ailandscape.org/tool/kokoro): Lightweight, high-quality open TTS model
- [MusicGen](https://ailandscape.org/tool/musicgen): Meta's controllable text-to-music generation model
- [XTTS](https://ailandscape.org/tool/xtts): Coqui's multilingual voice cloning TTS model

---

Part of [AI Landscape](https://ailandscape.org), an open map of the AI ecosystem. Web page: https://ailandscape.org/tool/styletts2 · Index for AI assistants: https://ailandscape.org/llms.txt
