Research graph
References from A dual-domain guided emotion-specialized swin transformer for enhancing speech emotion recognition. Local targets link to admitted publications; unresolved targets remain external evidence.
Feature pooling of modulation spectrum features for improved speech emotion recognition in the wild
10.1109/taffc.2018.2858255 · 2021 · External reference
PulseEmoNet: Pulse emotion network for speech emotion recognition
10.1016/j.bspc.2025.107687 · 2025 · External reference
Improving speech emotion recognition with adversarial data augmentation network
10.1109/tnnls.2020.3027600 · 2022 · External reference
Exploiting the potentialities of features for speech emotion recognition
10.1016/j.ins.2020.09.047 · 2021 · External reference
Temporal-Frequency state space duality: An efficient paradigm for speech emotion recognition
2025 · External reference
EIJL: Popularity prediction of social media advertisements based on multimodal emotional interaction and joint learning
10.3724/2096-7004.di.2025.0066 · 2025 · External reference
Speech emotion recognition based on convolutional neural network with Attention-Based bidirectional long short-term memory network and Multi-Task learning
10.1016/j.apacoust.2022.109178 · 2023 · External reference
Multiscale-Multichannel feature extraction and classification through one-dimensional convolutional neural network for speech emotion recognition
10.1016/j.specom.2023.103010 · 2024 · External reference
Hybrid CNN-BiLSTM architecture with multiple attention mechanisms to enhance speech emotion recognition
2025 · External reference
TRUST-SER: On the trustworthiness of Fine-Tuning Pre-Trained speech embeddings for speech emotion recognition
2024 · External reference
Enhancing intermodal interaction for unified Vision-Language understanding and generation
10.3724/2096-7004.di.2025.0034 · 2025 · External reference
Speech Swin-Transformer: Exploring a hierarchical Transformer with shifted windows for speech emotion recognition
2024 · External reference
Combining a parallel 2D CNN with a Self-Attention dilated residual network for CTC-Based discrete speech emotion recognition
10.1016/j.neunet.2021.03.013 · 2021 · External reference
Pre-Attentive speech signal processing with adaptive routing for emotion recognition
10.1016/j.bspc.2025.108782 · 2026 · External reference
DSTCNet: Deep Spectro-Temporal-Channel attention network for speech emotion recognition
10.1109/tnnls.2023.3304516 · 2025 · External reference
Foundation model assisted automatic speech emotion recognition: Transcribing, annotating, and augmenting
2024 · External reference
Attention is all you need
2017 · External reference
Two-stage emotion recognition framework using CNN-Transformer architecture and speaker cues
10.1016/j.apacoust.2025.110963 · 2025 · External reference
Dual-TBNet: Improving the robustness of speech features via dual-Transformer-BiLSTM for speech emotion recognition
10.1109/taslp.2023.3282092 · 2023 · External reference
Hierarchical convolutional neural networks with Post-Attention for speech emotion recognition
10.1016/j.neucom.2024.128879 · 2025 · External reference
BAT: Block and token Self-Attention for speech emotion recognition
10.1016/j.neunet.2022.09.022 · 2022 · External reference
SpeechFormer: A hierarchical efficient framework incorporating the characteristics of speech
2022 · External reference
SpeechFormer++: A hierarchical efficient framework for paralinguistic speech processing
10.1109/taslp.2023.3235194 · 2023 · External reference
Learning local to global feature aggregation for speech emotion recognition
2023 · External reference
Speech emotion recognition via an attentive Time-Frequency neural network
10.1109/tcss.2022.3219825 · 2023 · External reference
DST: Deformable speech Transformer for emotion recognition
2023 · External reference
DWFormer: Dynamic window transformer for speech emotion recognition
2023 · External reference
Aspect-Guided Multi-Graph convolutional networks for Aspect-based sentiment analysis
10.3724/2096-7004.di.2024.0052 · 2024 · External reference
Neural collapse under Cross-Entropy loss
10.1016/j.acha.2021.12.011 · 2022 · External reference
IEMOCAP: Interactive emotional dyadic motion capture database
10.1007/s10579-008-9076-6 · 2008 · External reference
Unresolved reference
External reference
10.21437/interspeech.2005-446
10.21437/interspeech.2005-446 · External reference
Sparse temporal aware capsule network for robust speech emotion recognition
10.1016/j.engappai.2025.110060 · 2025 · External reference
CENN: Capsule-Enhanced neural network with innovative metrics for robust speech emotion recognition
10.1016/j.knosys.2024.112499 · 2024 · External reference
Cubic knowledge distillation for speech emotion recognition
2024 · External reference
Leveraging speech PTM, text LLM, and emotional TTS for speech emotion recognition
2024 · External reference
AMSER: Accelerate mobile speech emotion recognition with signal compression
2025 · External reference
Exploiting wavelet scattering transform & Squeeze-Excitation blocks with Cross-Modal attention for Multi-Modal emotion recognition
2025 · External reference
Multi-Stream convolution-recurrent neural networks based on attention mechanism fusion for speech emotion recognition
10.3390/e24081025 · 2022 · External reference
A discriminative feature representation method based on cascaded attention network with adversarial strategy for speech emotion recognition
10.1109/taslp.2023.3245401 · 2023 · External reference
Speech emotion recognition using XGBoost and CNN BLSTM with attention
2021 · External reference
MA-CapsNet-DA: Speech emotion recognition based on MA-CapsNet using data augmentation
10.1016/j.eswa.2023.122939 · 2024 · External reference
The application of capsule neural network based CNN for speech emotion recognition
2021 · External reference
Research on psychological counseling and personality analysis algorithm based on speech emotion
2020 · External reference
Two-Layer fuzzy multiple random forest for speech emotion recognition in Human-Robot interaction
10.1016/j.ins.2019.09.005 · 2020 · External reference
An ensemble 1D-CNN-LSTM-GRU model with data augmentation for speech emotion recognition
2023 · External reference
Learning speech emotion representations in the Quaternion domain
10.1109/taslp.2023.3250840 · 2023 · External reference
Improving speech emotion recognition in under-resourced languages via Speech-to-Speech translation with bootstrapping data selection
2025 · External reference
Learning Multi-Scale features for speech emotion recognition with connection attention mechanism
10.1016/j.eswa.2022.118943 · 2023 · External reference
IEMOCAP: Interactive emotional dyadic motion capture database
10.1007/s10579-008-9076-6 · ExternalCitation · doi-reference
Neural collapse under Cross-Entropy loss
10.1016/j.acha.2021.12.011 · ExternalCitation · doi-reference
Speech emotion recognition based on convolutional neural network with Attention-Based bidirectional long short-term memory network and Multi-Task learning
10.1016/j.apacoust.2022.109178 · ExternalCitation · doi-reference
Two-stage emotion recognition framework using CNN-Transformer architecture and speaker cues
10.1016/j.apacoust.2025.110963 · ExternalCitation · doi-reference
PulseEmoNet: Pulse emotion network for speech emotion recognition
10.1016/j.bspc.2025.107687 · ExternalCitation · doi-reference
Pre-Attentive speech signal processing with adaptive routing for emotion recognition
10.1016/j.bspc.2025.108782 · ExternalCitation · doi-reference
Sparse temporal aware capsule network for robust speech emotion recognition
10.1016/j.engappai.2025.110060 · ExternalCitation · doi-reference
Learning Multi-Scale features for speech emotion recognition with connection attention mechanism
10.1016/j.eswa.2022.118943 · ExternalCitation · doi-reference
MA-CapsNet-DA: Speech emotion recognition based on MA-CapsNet using data augmentation
10.1016/j.eswa.2023.122939 · ExternalCitation · doi-reference
Two-Layer fuzzy multiple random forest for speech emotion recognition in Human-Robot interaction
10.1016/j.ins.2019.09.005 · ExternalCitation · doi-reference
Exploiting the potentialities of features for speech emotion recognition
10.1016/j.ins.2020.09.047 · ExternalCitation · doi-reference
CENN: Capsule-Enhanced neural network with innovative metrics for robust speech emotion recognition
10.1016/j.knosys.2024.112499 · ExternalCitation · doi-reference
Hierarchical convolutional neural networks with Post-Attention for speech emotion recognition
10.1016/j.neucom.2024.128879 · ExternalCitation · doi-reference
Combining a parallel 2D CNN with a Self-Attention dilated residual network for CTC-Based discrete speech emotion recognition
10.1016/j.neunet.2021.03.013 · ExternalCitation · doi-reference
BAT: Block and token Self-Attention for speech emotion recognition
10.1016/j.neunet.2022.09.022 · ExternalCitation · doi-reference
Multiscale-Multichannel feature extraction and classification through one-dimensional convolutional neural network for speech emotion recognition
10.1016/j.specom.2023.103010 · ExternalCitation · doi-reference
Feature pooling of modulation spectrum features for improved speech emotion recognition in the wild
10.1109/taffc.2018.2858255 · ExternalCitation · doi-reference
SpeechFormer++: A hierarchical efficient framework for paralinguistic speech processing
10.1109/taslp.2023.3235194 · ExternalCitation · doi-reference
A discriminative feature representation method based on cascaded attention network with adversarial strategy for speech emotion recognition
10.1109/taslp.2023.3245401 · ExternalCitation · doi-reference
Learning speech emotion representations in the Quaternion domain
10.1109/taslp.2023.3250840 · ExternalCitation · doi-reference
Dual-TBNet: Improving the robustness of speech features via dual-Transformer-BiLSTM for speech emotion recognition
10.1109/taslp.2023.3282092 · ExternalCitation · doi-reference
Speech emotion recognition via an attentive Time-Frequency neural network
10.1109/tcss.2022.3219825 · ExternalCitation · doi-reference
Improving speech emotion recognition with adversarial data augmentation network
10.1109/tnnls.2020.3027600 · ExternalCitation · doi-reference
DSTCNet: Deep Spectro-Temporal-Channel attention network for speech emotion recognition
10.1109/tnnls.2023.3304516 · ExternalCitation · doi-reference
10.21437/interspeech.2005-446
10.21437/interspeech.2005-446 · ExternalCitation · doi-reference
Multi-Stream convolution-recurrent neural networks based on attention mechanism fusion for speech emotion recognition
10.3390/e24081025 · ExternalCitation · doi-reference
Aspect-Guided Multi-Graph convolutional networks for Aspect-based sentiment analysis
10.3724/2096-7004.di.2024.0052 · ExternalCitation · doi-reference
Enhancing intermodal interaction for unified Vision-Language understanding and generation
10.3724/2096-7004.di.2025.0034 · ExternalCitation · doi-reference
EIJL: Popularity prediction of social media advertisements based on multimodal emotional interaction and joint learning
10.3724/2096-7004.di.2025.0066 · ExternalCitation · doi-reference