Parallel Network Speech Emotion Recognition Based on Hybrid Attention Mechanism.

Saved in:
Bibliographic Details
Title: Parallel Network Speech Emotion Recognition Based on Hybrid Attention Mechanism.
Authors: Hu, Zhangfang1 huzf@cqupt.edu.cn, Wang, Yulong2 2423245845@qq.com, Tang, Yicheng2 1035937714@qq.com
Source: IAENG International Journal of Computer Science. Jul2026, Vol. 53 Issue 7, p2740-2749. 10p.
Subjects: Emotion recognition, Parallel processing, Long short-term memory
Abstract: In speech emotion recognition tasks, relying on a single feature often limits the representational capacity for capturing emotional information, and using a single network model may lead to insufficient deep feature extraction, both of which can result in low classification accuracy. To address these issues, this paper proposes a parallel network architecture based on a hybrid attention mechanism, utilizing a combination of multiple features as network inputs to enhance emotion classification performance. The model first maps the 81-dimensional fused features into a 128-dimensional embedding space through an embedding layer, and then feeds them into three parallel subnetworks. Each subnetwork consists of an MSDC module, a bidirectional LSTM module, and a hybrid attention mechanism. Experimental results on the RAVDESS dataset show that the proposed method achieves an accuracy of 96.62% and a precision of 96.55% in an 8-class emotion classification task, outperforming other models and demonstrating the effectiveness and superiority of the proposed approach. [ABSTRACT FROM AUTHOR]
Copyright of IAENG International Journal of Computer Science is the property of International Association of Engineers (IAENG) and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract. (Copyright applies to all Abstracts.)
Database: Engineering Source
FullText Links:
  – Type: pdflink
Text:
  Availability: 0
Header DbId: egs
DbLabel: Engineering Source
An: 195088902
AccessLevel: 6
PubType: Academic Journal
PubTypeId: academicJournal
PreciseRelevancyScore: 0
IllustrationInfo
Items – Name: Title
  Label: Title
  Group: Ti
  Data: Parallel Network Speech Emotion Recognition Based on Hybrid Attention Mechanism.
– Name: Author
  Label: Authors
  Group: Au
  Data: <searchLink fieldCode="AR" term="%22Hu%2C+Zhangfang%22">Hu, Zhangfang</searchLink><relatesTo>1</relatesTo><i> huzf@cqupt.edu.cn</i><br /><searchLink fieldCode="AR" term="%22Wang%2C+Yulong%22">Wang, Yulong</searchLink><relatesTo>2</relatesTo><i> 2423245845@qq.com</i><br /><searchLink fieldCode="AR" term="%22Tang%2C+Yicheng%22">Tang, Yicheng</searchLink><relatesTo>2</relatesTo><i> 1035937714@qq.com</i>
– Name: TitleSource
  Label: Source
  Group: Src
  Data: <searchLink fieldCode="JN" term="%22IAENG+International+Journal+of+Computer+Science%22">IAENG International Journal of Computer Science</searchLink>. Jul2026, Vol. 53 Issue 7, p2740-2749. 10p.
– Name: Subject
  Label: Subjects
  Group: Su
  Data: <searchLink fieldCode="DE" term="%22Emotion+recognition%22">Emotion recognition</searchLink><br /><searchLink fieldCode="DE" term="%22Parallel+processing%22">Parallel processing</searchLink><br /><searchLink fieldCode="DE" term="%22Long+short-term+memory%22">Long short-term memory</searchLink>
– Name: Abstract
  Label: Abstract
  Group: Ab
  Data: In speech emotion recognition tasks, relying on a single feature often limits the representational capacity for capturing emotional information, and using a single network model may lead to insufficient deep feature extraction, both of which can result in low classification accuracy. To address these issues, this paper proposes a parallel network architecture based on a hybrid attention mechanism, utilizing a combination of multiple features as network inputs to enhance emotion classification performance. The model first maps the 81-dimensional fused features into a 128-dimensional embedding space through an embedding layer, and then feeds them into three parallel subnetworks. Each subnetwork consists of an MSDC module, a bidirectional LSTM module, and a hybrid attention mechanism. Experimental results on the RAVDESS dataset show that the proposed method achieves an accuracy of 96.62% and a precision of 96.55% in an 8-class emotion classification task, outperforming other models and demonstrating the effectiveness and superiority of the proposed approach. [ABSTRACT FROM AUTHOR]
– Name: AbstractSuppliedCopyright
  Label:
  Group: Ab
  Data: <i>Copyright of IAENG International Journal of Computer Science is the property of International Association of Engineers (IAENG) and its content may not be copied or emailed to multiple sites without the copyright holder's express written permission. Additionally, content may not be used with any artificial intelligence tools or machine learning technologies. However, users may print, download, or email articles for individual use. This abstract may be abridged. No warranty is given about the accuracy of the copy. Users should refer to the original published version of the material for the full abstract.</i> (Copyright applies to all Abstracts.)
PLink https://search.ebscohost.com/login.aspx?direct=true&site=eds-live&db=egs&AN=195088902
RecordInfo BibRecord:
  BibEntity:
    Languages:
      – Code: eng
        Text: English
    PhysicalDescription:
      Pagination:
        PageCount: 10
        StartPage: 2740
    Subjects:
      – SubjectFull: Emotion recognition
        Type: general
      – SubjectFull: Parallel processing
        Type: general
      – SubjectFull: Long short-term memory
        Type: general
    Titles:
      – TitleFull: Parallel Network Speech Emotion Recognition Based on Hybrid Attention Mechanism.
        Type: main
  BibRelationships:
    HasContributorRelationships:
      – PersonEntity:
          Name:
            NameFull: Hu, Zhangfang
      – PersonEntity:
          Name:
            NameFull: Wang, Yulong
      – PersonEntity:
          Name:
            NameFull: Tang, Yicheng
    IsPartOfRelationships:
      – BibEntity:
          Dates:
            – D: 01
              M: 07
              Text: Jul2026
              Type: published
              Y: 2026
          Identifiers:
            – Type: issn-print
              Value: 1819656X
          Numbering:
            – Type: volume
              Value: 53
            – Type: issue
              Value: 7
          Titles:
            – TitleFull: IAENG International Journal of Computer Science
              Type: main
ResultId 1