Código QR (código de barras bidimensional)

Multilingual speech-to-vocal tract visualization using deep learning for pronunciation training

Abstract Visualizing the vocal tract during speech remains a challenging task, even with recent advancements in open-source algorithms and datasets. A key limitation is the lack of multimodal resources that integrate audio with internal articulatory structures, which poses challenges to the developm...

תיאור מלא

שמור ב:
מידע ביבליוגרפי
Principais autores: Rodrigo Picinini Méxas, Yunji Chu, Unsang Park
פורמט: Artigo
שפה:Inglês
יצא לאור: SpringerOpen 2025-12-01
סדרה:EURASIP Journal on Audio, Speech, and Music Processing
נושאים:
גישה מקוונת:https://doi.org/10.1186/s13636-025-00438-x
תגים: הוספת תג
אין תגיות, היה/י הראשונ/ה לתייג את הרשומה!