Research on Dynamic Visemes of Japanese Speech Utterances
Research on Dynamic Visemes of Japanese Speech Utterances
批准号:
14580431
负责人:
HIRAYAMA Makoto J.
金额:
$2.62万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (C)
财政年份:
2002
资助国家:
日本
项目状态:
已结题
起止时间:
2002 至 2004
中文摘要
为了动画逼真的嘴唇运动在日本的讲话话语,高速视频录制完成,而一个人的主题说出简短的日语句子。两个高速摄像机捕捉正面和侧面的意见,人类主体的嘴唇在语音平衡的日语短句的讲话。记录速度为每秒300或240个样本,200 × 200或256 × 256像素,24位RGB颜色。在44.1 kHz Mono下记录声信号。除了自然嘴唇之外,对于一些句子记录嘴唇上的标记照明点和嘴唇上的蓝色着色,以使得能够更容易地建模提取嘴唇位置参数。这些录像被编译成日语视位视频数据库,这些视频图像被用来根据MPEG-4人体和人脸参数制作唇点的计算机图形模型参数。特别是对/ba/等爆炸性音素的快速动作进行逐帧观察,实现了嘴唇运动的精确再现,并将这些视位放置在时间轴上的关键帧上,在关键帧之间插入中间帧,完成了计算机图形动画。不仅是静态视位,而且还有动态视位,即,使用一组用于一个视位的多个帧的唇部参数,并使用非线性插值。使用这种方法,更真实的动画嘴唇运动在讲话中实现了比传统的简单插值相对较少的视位。
英文摘要
To animate realistic lip motions during Japanese speech utterances, high-speed video recording was done while a human subject uttered short Japanese sentences. Two high-speed video cameras captured front and side views of the human subject lips during speech utterances of phonemically balanced short Japanese sentences. Recordings were at 300 or 240 samples per second, 200 by 200 or 256 by 256 pixels, 24 bit RGB colors. Acoustic signals were recorded at 44.1 kHz Mono. Other than natural lips, marking illuminated points on lips and blue coloring on lips were recorded for some sentences to enable easier modeling of extracting lip position parameters. These recordings were compiled as a Japanese viseme video database.These video images were used to make computer graphics model parameters of lip points according to MPEG-4 body and face parameters. Especially, fast movements like explosive phonemes like /ba/ were carefully observed frame by frame to realize precise reproductions of lip movements.By placing these visemes onto key frames on time axis and interpolating middle frames between key frames, a computer graphics animation was done. Not only static visemes but also dynamic visemes, i.e., a set of lip parameters for multiple frames for one viseme, were used and nonlinear interpolations were used. Using this method, more realistic animations of lip motions during speech were realized than traditional simple interpolations of relatively small numbers of visemes.
期刊论文(20)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Making of a Japanese Viseme Video Database by Multiple High-speed Video Observations
通过多个高速视频观察制作日语视素视频数据库
DOI:
--
发表时间:
2003
期刊:
15th International Congress of Phonetic Sciences
影响因子:
--
作者:
[Makoto J.Hirayama]
通讯作者:
Makoto J.Hirayama
日本語音声発話口形素のCGモデル
日语口语视位的CG模型
DOI:
--
发表时间:
2004
期刊:
日本音響学会2004年秋季研究発表会講演論文集
影响因子:
--
作者:
[平山 亮]
通讯作者:
平山 亮
Makoto J.Hirayama: "Making of a Japanese Viseme Video Database by Multiple High-speed Video Observations"15th International Congress of Phonetic Sciences. 3157-3160 (2003)
Makoto J.Hirayama:“通过多个高速视频观察制作日语视素视频数据库”第 15 届国际语音科学大会。
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
A computer graphics model of Japanese visemes
日本发音嘴型的计算机图形模型
DOI:
--
发表时间:
2004
期刊:
2004 Autumn Meeting of Acoustical Society of Japan
影响因子:
--
作者:
[]
通讯作者:
平山 亮: "音声発話時の口唇周辺高速動画データベース"情報科学フォーラム(FIT)2003. 257-258 (2003)
Ryo Hirayama:“发声时嘴唇周围的高速视频数据库”信息科学论坛 (FIT) 2003. 257-258 (2003)
DOI:
--
发表时间:
期刊:
影响因子:
--
作者:
[]
通讯作者:
共 8 条
A research on tongue figure models of speech articulations
-
批准号:24500220
-
项目类别:Grant-in-Aid for Scientific Research (C)
-
资助金额:$3.41万
-
财政年份:2012
-
负责人:HIRAYAMA Makoto J.
-
依托单位: