QR kód

W2VC: WavLM representation based one-shot voice conversion with gradient reversal distillation and CTC supervision

Abstract Non-parallel data voice conversion (VC) has achieved considerable breakthroughs due to self-supervised pre-trained representation (SSPR) being used in recent years. Features extracted by the pre-trained model are expected to contain more content information. However, in common VC with SSPR,...

Celý popis

Uloženo v:
Podrobná bibliografie
Hlavní autoři: Hao Huang, Lin Wang, Jichen Yang, Ying Hu, Liang He
Médium: Artigo
Jazyk:Inglês
Vydáno: SpringerOpen 2023-10-01
Edice:EURASIP Journal on Audio, Speech, and Music Processing
Témata:
On-line přístup:https://doi.org/10.1186/s13636-023-00312-8
Tagy: Přidat tag
Žádné tagy, Buďte první, kdo vytvoří štítek k tomuto záznamu!