DSpace at KOASAS: Robust audio-visual speech recognition based on late integration

DSpace at KOASAS

College of Engineering(공과대학)School of Electrical Engineering(전기및전자공학부)EE-Journal Papers(저널논문)

Robust audio-visual speech recognition based on late integration

Cited 43 time in

Cited 0 time in

Hit : 305
Download : 0

Export

DC Field	Value	Language
dc.contributor.author	Lee, Jong-Seok	ko
dc.contributor.author	Park, Cheol Hoon	ko
dc.date.accessioned	2013-03-06T17:29:37Z	-
dc.date.available	2013-03-06T17:29:37Z	-
dc.date.created	2012-02-06	-
dc.date.created	2012-02-06	-
dc.date.issued	2008-08	-
dc.identifier.citation	IEEE TRANSACTIONS ON MULTIMEDIA, v.10, no.5, pp.767 - 779	-
dc.identifier.issn	1520-9210	-
dc.identifier.uri	http://hdl.handle.net/10203/87783	-
dc.description.abstract	Audio-visual speech recognition (AVSR) using acoustic and visual signals of speech has received attention because of its robustness in noisy environments. In this paper, we present a late integration scheme-based AVSR system whose robustness under various noise conditions is improved by enhancing the performance of the three parts composing the system. First, we improve the performance of the visual subsystem by using the stochastic optimization method for the hidden Markov models as the speech recognizer. Second, we propose a new method of considering dynamic characteristics of speech for improved robustness of the acoustic subsystem. Third, the acoustic and the visual subsystems are effectively integrated to produce final robust recognition results by using neural networks. We demonstrate the performance of the proposed methods via speaker-independent isolated word recognition experiments. The results show that the proposed system improves robustness over the conventional system under various noise conditions without a priori knowledge about the noise contained in the speech.	-
dc.language	English	-
dc.publisher	IEEE-Inst Electrical Electronics Engineers Inc	-
dc.subject	FUSION	-
dc.title	Robust audio-visual speech recognition based on late integration	-
dc.type	Article	-
dc.identifier.wosid	000258223800010	-
dc.identifier.scopusid	2-s2.0-47649103796	-
dc.type.rims	ART	-
dc.citation.volume	10	-
dc.citation.issue	5	-
dc.citation.beginningpage	767	-
dc.citation.endingpage	779	-
dc.citation.publicationname	IEEE TRANSACTIONS ON MULTIMEDIA	-
dc.identifier.doi	10.1109/TMM.2008.922789	-
dc.contributor.localauthor	Park, Cheol Hoon	-
dc.type.journalArticle	Article	-
dc.subject.keywordAuthor	audio-visual speech recognition	-
dc.subject.keywordAuthor	late integration	-
dc.subject.keywordAuthor	robustness hidden Markov model	-
dc.subject.keywordAuthor	interframe correlation	-
dc.subject.keywordAuthor	neural network	-
dc.subject.keywordAuthor	stochastic optimization	-
dc.subject.keywordPlus	FUSION	-

Appears in Collection: EE-Journal Papers(저널논문)

Files in This Item: There are no files associated with this item.

This item is cited by other documents in WoS

⊙ Detail Information in WoSⓡ	Click to see
⊙ Cited 43 items in WoS	Click to see citing articles in

Display Simple Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

Robust audio-visual speech recognition based on late integration

This item is cited by other documents in WoS

KOASAS

Communities & Collections