Speaker-Characterized Emotion Recognition using Online and Iterative Speaker Adaptation

Cited 8 time in webofscience Cited 0 time in scopus
  • Hit : 593
  • Download : 0
DC FieldValueLanguage
dc.contributor.authorKim, Jae-Bokko
dc.contributor.authorPark, Jeong-Sikko
dc.contributor.authorOh, Yung-Hwanko
dc.date.accessioned2013-03-12T14:02:27Z-
dc.date.available2013-03-12T14:02:27Z-
dc.date.created2013-01-21-
dc.date.created2013-01-21-
dc.date.issued2012-12-
dc.identifier.citationCOGNITIVE COMPUTATION, v.4, no.4, pp.398 - 408-
dc.identifier.issn1866-9956-
dc.identifier.urihttp://hdl.handle.net/10203/102532-
dc.description.abstractThis paper proposes a novel speech emotion recognition (SER) framework for affective interaction between human and personal devices. Most of the conventional SER techniques adopt a speaker-independent model framework because of the sparseness of individual speech data. However, a large amount of individual data can be accumulated on a personal device, making it possible to construct speaker-characterized emotion models in accordance with a speaker adaptation procedure. In this study, to address problems associated with conventional adaptation approaches in SER tasks, we modified a representative adaptation technique, maximum likelihood linear regression (MLLR), on the basis of selective label refinement. We subsequently carried out the modified MLLR procedure in an online and iterative manner, using accumulated individual data, to further enhance the speaker-characterized emotion models. In the SER experiments based on an emotional corpus, our approach exhibited performance superior to that of conventional adaptation techniques as well as the speaker-independent model framework.-
dc.languageEnglish-
dc.publisherSPRINGER-
dc.subjectHIDDEN MARKOV-MODELS-
dc.subjectSPEECH RECOGNITION-
dc.titleSpeaker-Characterized Emotion Recognition using Online and Iterative Speaker Adaptation-
dc.typeArticle-
dc.identifier.wosid000312126300003-
dc.identifier.scopusid2-s2.0-84870448053-
dc.type.rimsART-
dc.citation.volume4-
dc.citation.issue4-
dc.citation.beginningpage398-
dc.citation.endingpage408-
dc.citation.publicationnameCOGNITIVE COMPUTATION-
dc.identifier.doi10.1007/s12559-012-9132-9-
dc.contributor.localauthorOh, Yung-Hwan-
dc.contributor.nonIdAuthorKim, Jae-Bok-
dc.contributor.nonIdAuthorPark, Jeong-Sik-
dc.type.journalArticleArticle-
dc.subject.keywordAuthorSpeech emotion recognition-
dc.subject.keywordAuthorSpeaker adaptation-
dc.subject.keywordAuthorMaximum likelihood linear regression (MLLR)-
dc.subject.keywordAuthorSpeaker-characterized emotion model-
dc.subject.keywordAuthorHuman-machine interaction-
dc.subject.keywordPlusHIDDEN MARKOV-MODELS-
dc.subject.keywordPlusSPEECH RECOGNITION-
Appears in Collection
CS-Journal Papers(저널논문)
Files in This Item
There are no files associated with this item.
This item is cited by other documents in WoS
⊙ Detail Information in WoSⓡ Click to see webofscience_button
⊙ Cited 8 items in WoS Click to see citing articles in records_button

qr_code

  • mendeley

    citeulike


rss_1.0 rss_2.0 atom_1.0