This page cannot be shown here. You can still read it on the original site — the toolbar below keeps your place in the directory.
🐸 Github 🤗 Demo 🤖 Model card 💬 Discord We recently released XTTSv2 with 🐸TTS v0.20, and here I go over the relevant details of the model. XTTSv2 uses the same backbone as XTTSv1. It is a GPT2 model that predicts audio tokens computed by a pre-trained Discrete VAE model. The core update is changing the way we condition the model on the speaker information with a Perceiver model. In our model,…
Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.