The Voice AI Index / Text-to-Speech / #174
yl4579/StyleTTS2
by yl4579 · Text-to-Speech · updated 2y ago
StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
41
momentum
6,350
stars
703
forks
#174
rank
adversarial-trainingdeep-learningdiffusion-modelsganlatent-diffusionlatent-diffusion-modelspytorchspeaker-adaptationspeech-synthesistext-to-speechttswavlm
View on GitHub →