The Voice AI Index / Text-to-Speech / #174
yl4579

yl4579/StyleTTS2

by yl4579 · Text-to-Speech · updated 2y ago

StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models

41
momentum
6,350
stars
703
forks
#174
rank
adversarial-trainingdeep-learningdiffusion-modelsganlatent-diffusionlatent-diffusion-modelspytorchspeaker-adaptationspeech-synthesistext-to-speechttswavlm
View on GitHub →