Skip to main navigation Skip to search Skip to main content

Speech emotion recognition using eigen-FFT in clean and noisy environments

  • Ho Kim Eun
  • , Hak Hyun Kyung
  • , Hyun Kim Soo
  • , Keun Kwak Yoon
  • Korea Advanced Institute of Science and Technology

Research output: Contribution to conferenceConference paperpeer-review

Abstract

The ability to recognize human emotion is one of the basic techniques of human-robot interaction especially on emotional interaction points of view. Hence the purpose of this paper is to describe the realization of the recognition of emotion from a voice. Especially, this paper describes a speech emotion recognition technique using eigen-FFT. In the field of speech emotion recognition, recognition rate is important not only in clean but also in noisy environments. Hence, we constructed voice noise (or babble noise) data and performed recognition tests both in clean and noisy environments using the eigen-FFT. From the experiments, we achieved accuracy of 90.1±7.7% for four emotions at a 95% confidence interval and compared eigen-FFT with LPC, MFCC, and pitch in the clean environment. In the case of the noisy environments, eigen-FFT displayed superior performance over LPC, MFCC, and pitch.

Original languageEnglish
Title of host publication16th IEEE International Conference on Robot and Human Interactive Communication, RO-MAN
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages689-694
Number of pages6
ISBN (Print)1424416345, 9781424416349
DOIs
StatePublished - 2007
Event16th IEEE International Conference on Robot and Human Interactive Communication, RO-MAN 2007 - Jeju, Korea, Republic of
Duration: 2007.08.262007.08.29

Publication series

NameProceedings - IEEE International Workshop on Robot and Human Interactive Communication

Conference

Conference16th IEEE International Conference on Robot and Human Interactive Communication, RO-MAN 2007
Country/TerritoryKorea, Republic of
CityJeju
Period07.08.2607.08.29

Fingerprint

Dive into the research topics of 'Speech emotion recognition using eigen-FFT in clean and noisy environments'. Together they form a unique fingerprint.

Cite this