kod QR

EmoBridge: Aligning Speech and Language for Emotion Recognition via Q-Former

Speech emotion recognition (SER) remains a challenging task due to the limited affective cues in unimodal representations and the difficulty of aligning heterogeneous features in multimodal systems. Although multimodal large language models (MLLMs) have recently achieved notable progress in affectiv...

Szczegółowa specyfikacja

Zapisane w:
Opis bibliograficzny
Główni autorzy: Yuntao Sun, Xuehua Zhang, Yanting Sun
Format: Artigo
Język:Inglês
Wydane: IEEE 2025-01-01
Seria:IEEE Access
Hasła przedmiotowe:
Dostęp online:https://ieeexplore.ieee.org/document/11263794/
Etykiety: Dodaj etykietę
Nie ma etykietki, Dołącz pierwszą etykiete!