Logo image
Speech in Affective Computing
Book chapter

Speech in Affective Computing

Chi-Chun Lee, Jangwon Kim, Angeliki Metallinou, Carlos Busso, Sungbok Lee and Shrikanth S. Narayanan
Oxford Handbooks Online The Oxford Handbook of Affective Computing
2015

Abstract

emotional speech production;acoustic feature extraction for emotion analysis;computational frameworks for emotion recognition;acoustic feature normalization
This chapter is from the forthcoming The Oxford Handbook of Affective Computing edited by Rafael Calvo, Sidney K. D'Mello, Jonathan Gratch, and Arvid Kappas. Speech is a key communication modality for humans to encode emotion. In this chapter, we address three main aspects of speech in affective computing: emotional speech production, acoustic feature extraction for emotion analysis, and the design of a speech-based emotion recognizer. Specifically we discuss the current understanding of the interplay of speech production vocal organs during expressive speech, extracting informative acoustic features from speech recording waveforms, and the engineering design of automatic emotion recognizers using speech acoustic-based features. The latter includes a discussion of emotion labeling for generating ground truth references, acoustic feature normalization for controlling signal variability, and choice of computational frameworks for emotion recognition. Finally, we present some open challenges and applications of a robust emotion recognizer.

Metrics

1 Record Views

Details

Logo image