Advancing Bangla text-to-speech synthesis using a VITS-based model with a custom dataset and comprehensive evaluation
Abstract Bangla (Bengali) voice synthesis presents distinct issues due to its restricted language resources and complex vocabulary. This work introduces a state-of-the-art Bangla text-to-speech (TTS) system that utilizes the Variational Inference Synthesis of Speech (VITS) architecture. We tailor VI...
Guardat en:
| Autors principals: | , , , |
|---|---|
| Format: | Artigo |
| Idioma: | Inglês |
| Publicat: |
Springer
2025-08-01
|
| Col·lecció: | Discover Computing |
| Matèries: | |
| Accés en línia: | https://doi.org/10.1007/s10791-025-09701-3 |
| Etiquetes: |
Sense etiquetes, Sigues el primer a etiquetar aquest registre!
|
