W2VC: WavLM representation based one-shot voice conversion with gradient reversal distillation and CTC supervision
Abstract Non-parallel data voice conversion (VC) has achieved considerable breakthroughs due to self-supervised pre-trained representation (SSPR) being used in recent years. Features extracted by the pre-trained model are expected to contain more content information. However, in common VC with SSPR,...
Na minha lista:
| Principais autores: | , , , , |
|---|---|
| Formato: | Artigo |
| Idioma: | Inglês |
| Publicado em: |
SpringerOpen
2023-10-01
|
| coleção: | EURASIP Journal on Audio, Speech, and Music Processing |
| Assuntos: | |
| Acesso em linha: | https://doi.org/10.1186/s13636-023-00312-8 |
| Tags: |
Sem tags, seja o primeiro a adicionar uma tag!
|
