Speech Emotion Recognition Using Multi-Scale Global–Local Representation Learning with Feature Pyramid Network
Speech emotion recognition (SER) is important in facilitating natural human–computer interactions. In speech sequence modeling, a vital challenge is to learn context-aware sentence expression and temporal dynamics of paralinguistic features to achieve unambiguous emotional semantic understanding. In...
保存先:
| 主要な著者: | , , , , |
|---|---|
| フォーマット: | Artigo |
| 言語: | Inglês |
| 出版事項: |
MDPI AG
2024-12-01
|
| シリーズ: | Applied Sciences |
| 主題: | |
| オンライン・アクセス: | https://www.mdpi.com/2076-3417/14/24/11494 |
| タグ: |
タグなし, このレコードへの初めてのタグを付けませんか!
|
