Abstract
Contact and support
Need help, have a question, or want to contact the ResearchHub team?
© 2026 ResearchHub. Built for responsible scholarly connection.
Yonghong Fan, Heming Huang, Feipeng Da
Abstract
Authors
Institutions
Provenance
crossref
Confidence 100%
openalex
Confidence 95%
No local reference links have been materialized yet.
No local citing links have been materialized yet.
Feature pooling of modulation spectrum features for improved speech emotion recognition in the wild
10.1109/taffc.2018.2858255 · 2021
PulseEmoNet: Pulse emotion network for speech emotion recognition
10.1016/j.bspc.2025.107687 · 2025
Improving speech emotion recognition with adversarial data augmentation network
10.1109/tnnls.2020.3027600 · 2022
Exploiting the potentialities of features for speech emotion recognition
10.1016/j.ins.2020.09.047 · 2021
Temporal-Frequency state space duality: An efficient paradigm for speech emotion recognition
2025
EIJL: Popularity prediction of social media advertisements based on multimodal emotional interaction and joint learning
10.3724/2096-7004.di.2025.0066 · 2025
Speech emotion recognition based on convolutional neural network with Attention-Based bidirectional long short-term memory network and Multi-Task learning
10.1016/j.apacoust.2022.109178 · 2023
Multiscale-Multichannel feature extraction and classification through one-dimensional convolutional neural network for speech emotion recognition
10.1016/j.specom.2023.103010 · 2024
Hybrid CNN-BiLSTM architecture with multiple attention mechanisms to enhance speech emotion recognition
datacite
Confidence 0%
2025
TRUST-SER: On the trustworthiness of Fine-Tuning Pre-Trained speech embeddings for speech emotion recognition
2024
Enhancing intermodal interaction for unified Vision-Language understanding and generation
10.3724/2096-7004.di.2025.0034 · 2025
Speech Swin-Transformer: Exploring a hierarchical Transformer with shifted windows for speech emotion recognition
2024
Combining a parallel 2D CNN with a Self-Attention dilated residual network for CTC-Based discrete speech emotion recognition
10.1016/j.neunet.2021.03.013 · 2021
Pre-Attentive speech signal processing with adaptive routing for emotion recognition
10.1016/j.bspc.2025.108782 · 2026
DSTCNet: Deep Spectro-Temporal-Channel attention network for speech emotion recognition
10.1109/tnnls.2023.3304516 · 2025
Foundation model assisted automatic speech emotion recognition: Transcribing, annotating, and augmenting
2024
Attention is all you need
2017
Two-stage emotion recognition framework using CNN-Transformer architecture and speaker cues
10.1016/j.apacoust.2025.110963 · 2025
Dual-TBNet: Improving the robustness of speech features via dual-Transformer-BiLSTM for speech emotion recognition
10.1109/taslp.2023.3282092 · 2023
Hierarchical convolutional neural networks with Post-Attention for speech emotion recognition
10.1016/j.neucom.2024.128879 · 2025
BAT: Block and token Self-Attention for speech emotion recognition
10.1016/j.neunet.2022.09.022 · 2022
SpeechFormer: A hierarchical efficient framework incorporating the characteristics of speech
2022
SpeechFormer++: A hierarchical efficient framework for paralinguistic speech processing
10.1109/taslp.2023.3235194 · 2023
Learning local to global feature aggregation for speech emotion recognition
2023
Speech emotion recognition via an attentive Time-Frequency neural network
10.1109/tcss.2022.3219825 · 2023
DST: Deformable speech Transformer for emotion recognition
2023
DWFormer: Dynamic window transformer for speech emotion recognition
2023
Aspect-Guided Multi-Graph convolutional networks for Aspect-based sentiment analysis
10.3724/2096-7004.di.2024.0052 · 2024
Neural collapse under Cross-Entropy loss
10.1016/j.acha.2021.12.011 · 2022
IEMOCAP: Interactive emotional dyadic motion capture database
10.1007/s10579-008-9076-6 · 2008
Unresolved referenced work
Kept as external metadata until matched
10.21437/interspeech.2005-446
10.21437/interspeech.2005-446
Sparse temporal aware capsule network for robust speech emotion recognition
10.1016/j.engappai.2025.110060 · 2025
CENN: Capsule-Enhanced neural network with innovative metrics for robust speech emotion recognition
10.1016/j.knosys.2024.112499 · 2024
Cubic knowledge distillation for speech emotion recognition
2024
Leveraging speech PTM, text LLM, and emotional TTS for speech emotion recognition
2024
AMSER: Accelerate mobile speech emotion recognition with signal compression
2025
Exploiting wavelet scattering transform & Squeeze-Excitation blocks with Cross-Modal attention for Multi-Modal emotion recognition
2025
Multi-Stream convolution-recurrent neural networks based on attention mechanism fusion for speech emotion recognition
10.3390/e24081025 · 2022
A discriminative feature representation method based on cascaded attention network with adversarial strategy for speech emotion recognition
10.1109/taslp.2023.3245401 · 2023
Learning Multi-Scale features for speech emotion recognition with connection attention mechanism
10.1016/j.eswa.2022.118943 · doi-reference
Learning speech emotion representations in the Quaternion domain
10.1109/taslp.2023.3250840 · doi-reference
Two-Layer fuzzy multiple random forest for speech emotion recognition in Human-Robot interaction
10.1016/j.ins.2019.09.005 · doi-reference
MA-CapsNet-DA: Speech emotion recognition based on MA-CapsNet using data augmentation
10.1016/j.eswa.2023.122939 · doi-reference
A discriminative feature representation method based on cascaded attention network with adversarial strategy for speech emotion recognition
10.1109/taslp.2023.3245401 · doi-reference
Multi-Stream convolution-recurrent neural networks based on attention mechanism fusion for speech emotion recognition
10.3390/e24081025 · doi-reference
CENN: Capsule-Enhanced neural network with innovative metrics for robust speech emotion recognition
10.1016/j.knosys.2024.112499 · doi-reference
Sparse temporal aware capsule network for robust speech emotion recognition
10.1016/j.engappai.2025.110060 · doi-reference
10.21437/interspeech.2005-446
10.21437/interspeech.2005-446 · doi-reference
IEMOCAP: Interactive emotional dyadic motion capture database
10.1007/s10579-008-9076-6 · doi-reference
Neural collapse under Cross-Entropy loss
10.1016/j.acha.2021.12.011 · doi-reference
Aspect-Guided Multi-Graph convolutional networks for Aspect-based sentiment analysis
10.3724/2096-7004.di.2024.0052 · doi-reference
Speech emotion recognition via an attentive Time-Frequency neural network
10.1109/tcss.2022.3219825 · doi-reference
SpeechFormer++: A hierarchical efficient framework for paralinguistic speech processing
10.1109/taslp.2023.3235194 · doi-reference
BAT: Block and token Self-Attention for speech emotion recognition
10.1016/j.neunet.2022.09.022 · doi-reference
Hierarchical convolutional neural networks with Post-Attention for speech emotion recognition
10.1016/j.neucom.2024.128879 · doi-reference
Dual-TBNet: Improving the robustness of speech features via dual-Transformer-BiLSTM for speech emotion recognition
10.1109/taslp.2023.3282092 · doi-reference
Two-stage emotion recognition framework using CNN-Transformer architecture and speaker cues
10.1016/j.apacoust.2025.110963 · doi-reference
DSTCNet: Deep Spectro-Temporal-Channel attention network for speech emotion recognition
10.1109/tnnls.2023.3304516 · doi-reference
Pre-Attentive speech signal processing with adaptive routing for emotion recognition
10.1016/j.bspc.2025.108782 · doi-reference
Combining a parallel 2D CNN with a Self-Attention dilated residual network for CTC-Based discrete speech emotion recognition
10.1016/j.neunet.2021.03.013 · doi-reference
Enhancing intermodal interaction for unified Vision-Language understanding and generation
10.3724/2096-7004.di.2025.0034 · doi-reference
Multiscale-Multichannel feature extraction and classification through one-dimensional convolutional neural network for speech emotion recognition
10.1016/j.specom.2023.103010 · doi-reference
Speech emotion recognition based on convolutional neural network with Attention-Based bidirectional long short-term memory network and Multi-Task learning
10.1016/j.apacoust.2022.109178 · doi-reference
EIJL: Popularity prediction of social media advertisements based on multimodal emotional interaction and joint learning
10.3724/2096-7004.di.2025.0066 · doi-reference
Exploiting the potentialities of features for speech emotion recognition
10.1016/j.ins.2020.09.047 · doi-reference
Improving speech emotion recognition with adversarial data augmentation network
10.1109/tnnls.2020.3027600 · doi-reference
PulseEmoNet: Pulse emotion network for speech emotion recognition
10.1016/j.bspc.2025.107687 · doi-reference
Feature pooling of modulation spectrum features for improved speech emotion recognition in the wild
10.1109/taffc.2018.2858255 · doi-reference