Codi QR

Audio–Visual Speech Recognition Based on Dual Cross-Modality Attentions with the Transformer Model

Since attention mechanism was introduced in neural machine translation, attention has been combined with the long short-term memory (LSTM) or replaced the LSTM in a transformer model to overcome the sequence-to-sequence (seq2seq) problems with the LSTM. In contrast to the neural machine translation,...

Descripció completa

Guardat en:
Dades bibliogràfiques
Autors principals: Yong-Hyeok Lee, Dong-Won Jang, Jae-Bin Kim, Rae-Hong Park, Hyung-Min Park
Format: Artigo
Idioma:Inglês
Publicat: MDPI AG 2020-10-01
Col·lecció:Applied Sciences
Matèries:
Accés en línia:https://www.mdpi.com/2076-3417/10/20/7263
Etiquetes: Afegir etiqueta
Sense etiquetes, Sigues el primer a etiquetar aquest registre!