Cross-feature fusion speech emotion recognition based on attention mask residual network and Wav2vec 2.0
Speech Emotion Recognition (SER) has received widespread attention as a crucial way for understanding human emotional states. However, the impact of irrelevant information on speech signals and data sparsity limit the development of SER system. To address these issues, this paper proposes a framewor...
Đã lưu trong:
| Những tác giả chính: | , |
|---|---|
| Định dạng: | Artigo |
| Ngôn ngữ: | Inglês |
| Được phát hành: |
KeAi Communications Co., Ltd.
2025-10-01
|
| Loạt: | Digital Communications and Networks |
| Những chủ đề: | |
| Truy cập trực tuyến: | http://www.sciencedirect.com/science/article/pii/S2352864824001299 |
| Các nhãn: |
Không có thẻ, Là người đầu tiên thẻ bản ghi này!
|
