Refereed conference paper presented and published in conference proceedings
香港中文大学研究人员 ( 现职)
蒙美玲教授 (系统工程与工程管理学系) |
卢伟杰博士 (系统工程与工程管理学系) |
梁伟俭先生 (系统工程与工程管理学系) |
苏帼慧教授 (那打素护理学院) |
黄嘉豪先生 (系统工程与工程管理学系) |
全文
数位物件识别号 (DOI) http://dx.doi.org/10.1109/ISCSLP.2010.5684832 |
引用次数
Scopushttp://aims.cuhk.edu.hk/converis/portal/Publication/5Scopus source URL
其它资讯
摘要This paper presents a two-dimensional (2D) visual-speech synthesizer to support language learning. A visual-speech synthesizer animates the human articulators in synchronization with speech signals, e.g., output from a text-to-speech synthesizer. A visual-speech animation can offer a concrete illustration to the language learners on how to move and where to place the articulators when pronouncing a phoneme. We adopt a 2D vector-based viseme models and compiled a collection of visemes to cover the articulation of all English phonemes (42 visemes for the 44 English phonemes). Morphing between properly selected vector-based articulation images achieves articulatory animations. In this way, we have developed an articulatory visual speech synthesizer that can accept free-text input and synthesize articulatory dynamics in real-time. Evaluation involving 32 subjects based on "lip-reading" shows that they can identify the appropriate word(s) based on articulation animation alone nearly ~80% of the time. ?2010 IEEE.
着者Wong K.-H., Leung W.-K., Lo W.-K., Meng H.
会议名称2010 7th International Symposium on Chinese Spoken Language Processing, ISCSLP 2010
会议开始日29.11.2010
会议完结日03.12.2010
会议地点Tainan
会议国家台湾
详细描述ISCA and IEEE
出版年份2010
月份12
日期1
页次139 - 143
国际标準书号9781424462469
语言英式英语
关键词Articulatory phonetics, Computer-assisted language learning, Text-to-audiovisual systhesis, Visual-speech synthesizer