Hierarchical modeling of temporal course in emotional expression for speech emotion recognition

Chung Hsien Wu, Wei Bin Liang, Kuan Chun Cheng, Jen Chun Lin

研究成果: Conference contribution

4 引文 (Scopus)

摘要

This paper presents an approach to hierarchical modeling of temporal course in emotional expression for speech emotion recognition. In the proposed approach, a segmentation algorithm is employed to hierarchically chunk an input utterance into three-level temporal units, including low-level descriptors (LLDs)-based sub-utterance level, emotion profile (EP)-based sub-utterance level and utterance level. An emotion-oriented hierarchical structure is constructed based on the three-level units to describe the temporal emotion expression in an utterance. A hierarchical correlation model is also proposed to fuse the three-level outputs from the corresponding emotion recognizers and further model the correlation among them to determine the emotional state of the utterance. The EMO-DB corpus was used to evaluate the performance on speech emotion recognition. Experimental results show that the proposed method considering the temporal course in emotional expression provides the potential to improve the speech emotion recognition performance.

原文English
主出版物標題2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015
發行者Institute of Electrical and Electronics Engineers Inc.
頁面810-814
頁數5
ISBN(電子)9781479999538
DOIs
出版狀態Published - 2015 十二月 2
事件2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015 - Xi'an, China
持續時間: 2015 九月 212015 九月 24

出版系列

名字2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015

Other

Other2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015
國家China
城市Xi'an
期間15-09-2115-09-24

指紋

Speech recognition
Electric fuses

All Science Journal Classification (ASJC) codes

  • Artificial Intelligence
  • Computer Vision and Pattern Recognition
  • Human-Computer Interaction
  • Software

引用此文

Wu, C. H., Liang, W. B., Cheng, K. C., & Lin, J. C. (2015). Hierarchical modeling of temporal course in emotional expression for speech emotion recognition. 於 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015 (頁 810-814). [7344666] (2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015). Institute of Electrical and Electronics Engineers Inc.. https://doi.org/10.1109/ACII.2015.7344666
Wu, Chung Hsien ; Liang, Wei Bin ; Cheng, Kuan Chun ; Lin, Jen Chun. / Hierarchical modeling of temporal course in emotional expression for speech emotion recognition. 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015. Institute of Electrical and Electronics Engineers Inc., 2015. 頁 810-814 (2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015).
@inproceedings{fceebc2c94f247d59db8460c9dfa7a31,
title = "Hierarchical modeling of temporal course in emotional expression for speech emotion recognition",
abstract = "This paper presents an approach to hierarchical modeling of temporal course in emotional expression for speech emotion recognition. In the proposed approach, a segmentation algorithm is employed to hierarchically chunk an input utterance into three-level temporal units, including low-level descriptors (LLDs)-based sub-utterance level, emotion profile (EP)-based sub-utterance level and utterance level. An emotion-oriented hierarchical structure is constructed based on the three-level units to describe the temporal emotion expression in an utterance. A hierarchical correlation model is also proposed to fuse the three-level outputs from the corresponding emotion recognizers and further model the correlation among them to determine the emotional state of the utterance. The EMO-DB corpus was used to evaluate the performance on speech emotion recognition. Experimental results show that the proposed method considering the temporal course in emotional expression provides the potential to improve the speech emotion recognition performance.",
author = "Wu, {Chung Hsien} and Liang, {Wei Bin} and Cheng, {Kuan Chun} and Lin, {Jen Chun}",
year = "2015",
month = "12",
day = "2",
doi = "10.1109/ACII.2015.7344666",
language = "English",
series = "2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015",
publisher = "Institute of Electrical and Electronics Engineers Inc.",
pages = "810--814",
booktitle = "2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015",
address = "United States",

}

Wu, CH, Liang, WB, Cheng, KC & Lin, JC 2015, Hierarchical modeling of temporal course in emotional expression for speech emotion recognition. 於 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015., 7344666, 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015, Institute of Electrical and Electronics Engineers Inc., 頁 810-814, 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015, Xi'an, China, 15-09-21. https://doi.org/10.1109/ACII.2015.7344666

Hierarchical modeling of temporal course in emotional expression for speech emotion recognition. / Wu, Chung Hsien; Liang, Wei Bin; Cheng, Kuan Chun; Lin, Jen Chun.

2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015. Institute of Electrical and Electronics Engineers Inc., 2015. p. 810-814 7344666 (2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015).

研究成果: Conference contribution

TY - GEN

T1 - Hierarchical modeling of temporal course in emotional expression for speech emotion recognition

AU - Wu, Chung Hsien

AU - Liang, Wei Bin

AU - Cheng, Kuan Chun

AU - Lin, Jen Chun

PY - 2015/12/2

Y1 - 2015/12/2

N2 - This paper presents an approach to hierarchical modeling of temporal course in emotional expression for speech emotion recognition. In the proposed approach, a segmentation algorithm is employed to hierarchically chunk an input utterance into three-level temporal units, including low-level descriptors (LLDs)-based sub-utterance level, emotion profile (EP)-based sub-utterance level and utterance level. An emotion-oriented hierarchical structure is constructed based on the three-level units to describe the temporal emotion expression in an utterance. A hierarchical correlation model is also proposed to fuse the three-level outputs from the corresponding emotion recognizers and further model the correlation among them to determine the emotional state of the utterance. The EMO-DB corpus was used to evaluate the performance on speech emotion recognition. Experimental results show that the proposed method considering the temporal course in emotional expression provides the potential to improve the speech emotion recognition performance.

AB - This paper presents an approach to hierarchical modeling of temporal course in emotional expression for speech emotion recognition. In the proposed approach, a segmentation algorithm is employed to hierarchically chunk an input utterance into three-level temporal units, including low-level descriptors (LLDs)-based sub-utterance level, emotion profile (EP)-based sub-utterance level and utterance level. An emotion-oriented hierarchical structure is constructed based on the three-level units to describe the temporal emotion expression in an utterance. A hierarchical correlation model is also proposed to fuse the three-level outputs from the corresponding emotion recognizers and further model the correlation among them to determine the emotional state of the utterance. The EMO-DB corpus was used to evaluate the performance on speech emotion recognition. Experimental results show that the proposed method considering the temporal course in emotional expression provides the potential to improve the speech emotion recognition performance.

UR - http://www.scopus.com/inward/record.url?scp=84964040014&partnerID=8YFLogxK

UR - http://www.scopus.com/inward/citedby.url?scp=84964040014&partnerID=8YFLogxK

U2 - 10.1109/ACII.2015.7344666

DO - 10.1109/ACII.2015.7344666

M3 - Conference contribution

AN - SCOPUS:84964040014

T3 - 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015

SP - 810

EP - 814

BT - 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015

PB - Institute of Electrical and Electronics Engineers Inc.

ER -

Wu CH, Liang WB, Cheng KC, Lin JC. Hierarchical modeling of temporal course in emotional expression for speech emotion recognition. 於 2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015. Institute of Electrical and Electronics Engineers Inc. 2015. p. 810-814. 7344666. (2015 International Conference on Affective Computing and Intelligent Interaction, ACII 2015). https://doi.org/10.1109/ACII.2015.7344666