Cross-feature fusion speech emotion recognition based on attention mask residual network and Wav2vec 2.0
Speech Emotion Recognition (SER) has received widespread attention as a crucial way for understanding human emotional states. However, the impact of irrelevant information on speech signals and data sparsity limit the development of SER system. To address these issues, this paper proposes a framewor...
Na minha lista:
| Principais autores: | , |
|---|---|
| 格式: | Artigo |
| 語言: | Inglês |
| 出版: |
KeAi Communications Co., Ltd.
2025-10-01
|
| 叢編: | Digital Communications and Networks |
| 主題: | |
| 在線閱讀: | http://www.sciencedirect.com/science/article/pii/S2352864824001299 |
| 標簽: |
沒有標簽, 成為第一個標記此記錄!
|
