Codi QR

Advancing Bangla text-to-speech synthesis using a VITS-based model with a custom dataset and comprehensive evaluation

Abstract Bangla (Bengali) voice synthesis presents distinct issues due to its restricted language resources and complex vocabulary. This work introduces a state-of-the-art Bangla text-to-speech (TTS) system that utilizes the Variational Inference Synthesis of Speech (VITS) architecture. We tailor VI...

Descripció completa

Guardat en:
Dades bibliogràfiques
Autors principals: Sujeet Kumar, Siddharth Kumar, Kushal Sathe, Jayadeep Pati
Format: Artigo
Idioma:Inglês
Publicat: Springer 2025-08-01
Col·lecció:Discover Computing
Matèries:
Accés en línia:https://doi.org/10.1007/s10791-025-09701-3
Etiquetes: Afegir etiqueta
Sense etiquetes, Sigues el primer a etiquetar aquest registre!