StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
StyleTTS 2 reaches near-human quality using style diffusion and adversarial training, and can clone a voice from a few seconds of audio.
git clone https://github.com/yl4579/StyleTTS2.git cd StyleTTS2
Excerpts from the project README on GitHub. Copyright and licensing remain with the respective authors.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeStyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
StyleTTS2 has 6.3k stars on GitHub. It has been forked 701 times. StyleTTS2 is written mainly in Python. It has been in active development since 2023. StyleTTS2 is available under the MIT license. Its main topics are adversarial-training, deep-learning, diffusion-models, gan.
Read the full guideStyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
StyleTTS2 is an open-source project. It is released under the MIT license.
Yes. StyleTTS2 is free and open source — you can use, modify and self-host it.
StyleTTS2 is available under the MIT license.
StyleTTS2 is written mainly in Python.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/yl4579-styletts2.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.