A Study on a framework of spontaneous communication depending on dialogue situation
A Study on a framework of spontaneous communication depending on dialogue situation
批准号:
17300066
负责人:
SHIRAI Katsuhiko
金额:
$9.83万
依托单位:
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2005
资助国家:
日本
项目状态:
已结题
起止时间:
2005 至 2007
中文摘要
在这项研究中,我们研究了一个与用户自发交互的通信系统框架,以应对实际的对话环境。传统的口语对话系统旨在有效地完成特定的语音对话任务,而自发交流系统的框架是为了将这些口语对话系统进化为自发地开始和继续口语对话的系统而开发的。为此,我们进行了三个方面的研究:(1)对对话环境的理解,旨在利用图像和语音信号进行高级人类识别和口语对话识别;(2)自发交流管理模型,为口语对话的开始、继续和结束建模;(3)语音生成和动作表达技术,即如何通过话语或动作来呈现系统的意图。在研究(i)中,开发了基于立体摄像机的人体姿态估计方法。同时利用图像的空间深度、人体或衣服的形状和纹理信息,实现了对人体姿态的准确估计。此外,还研究了话语意图的估计。利用句尾和单词n -gram的特征,实现了更准确的话语意图。在研究(ii)中,开发了机器人与人开始交流、继续交流和结束交流的模型。为此,我们开发了对话伙伴的心理状态。在研究中,开发了一种笑声语音生成技术。在对人类笑声进行声学分析的基础上,实现了笑声和笑语的合成。通过这三项研究,建立了自发交际的基本框架。
英文摘要
In this study, we examined a framework of communication systems that interact with users spontaneously so as to cope with practical dialogue environments. While conventional spoken dialogue systems aimed to efficiently achieve specific speech-dialogue tasks, the framework of spontaneous communication system was developed in order to evolve these spoken dialogue systems to one that spontaneously start and continue spoken dialogues.For this purpose, three studies were conducted : (i) understanding of dialogue environment, which aims advanced human recognition and spoken dialogue recognition using image and speech signal, (ii) spontaneous communication management model, which models how to start, continue and end spoken dialogues, and(iii) speech generation and motion expression technology, which is how to present intentions of the system by utterances or motions.In the study(i), human pose estimation using stereo camera was developed. Simultaneous adoption of information of space depth of images, and shapes and textures of either human bodies or clothes realized accurate estimation of human poses. In addition, estimation of utterance intention was studied. Using characteristics of end of sentences and word N-grams, more accurate utterance intention was achieved.In the study(ii), models for a robot to start communication with a human, to continue it, and to end it were developed. we developed mental-state of a conversational partner for these purposes.In the study a speech generation technique for laughter was developed. Based on acoustical analyses of human laughter syntheses of both laughter and laughter-speech were realized. Throughout these three studies, basic framework of spontaneous communication was established.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Fusion-based Age-group Classification Method Using Multiple Two-Dimensional Featore Extraction Algorithms
基于融合的多种二维特征提取算法的年龄组分类方法
DOI:
--
发表时间:
2007
期刊:
IEICE Trans. on Information and Systems(ED) Vol.E90-D, No.6
影响因子:
--
作者:
[Kazuya Ueki, Tetsunori Kobayashi]
通讯作者:
Tetsunori Kobayashi
SPECTRAL FREQUENCY TRACKING FOR CLASSIFYING AUDIO SIGNALS
用于对音频信号进行分类的频谱跟踪
DOI:
--
发表时间:
2006
期刊:
Proc. of 2006 IEEE International Symposium on Signal Processing and Information Technology
影响因子:
--
作者:
[Toru Taniguchi, Mikio Tohyama, Katsuhiko Shirai]
通讯作者:
Katsuhiko Shirai
Sinusoidal Segmentの時間的特徴を用いた音声・楽器音・歌声が混在した音響信号中の音カテゴリ検出
使用正弦片段的时间特征检测包含语音、乐器声音和歌声混合的声学信号中的声音类别
DOI:
--
发表时间:
2005
期刊:
日本音響学会秋季研究発表会講演論文集 2-6-5
影响因子:
--
作者:
[谷口徹, 安達了慈, 大川茂樹, 誉田雅彰, 白井克彦]
通讯作者:
白井克彦
言語獲得における母音範疇の形成過程のシミュレーション
语言习得中元音类别形成过程的模拟
DOI:
--
发表时间:
2007
期刊:
影响因子:
--
作者:
[宮澤 幸希, 白勢 彩子, 菊池 英明]
通讯作者:
菊池 英明
DOI:
--
发表时间:
2005
期刊:
影响因子:
--
作者:
[Kei Nakajima, Shinya Fujie, Yousuke Matsuzaka, Tetsunori Kobayashi]
通讯作者:
Tetsunori Kobayashi
共 75 条
Construction of Multimodal Emotion Representation Model for Computer Animation
-
批准号:14208031
-
项目类别:Grant-in-Aid for Scientific Research (A)
-
资助金额:$26.87万
-
财政年份:2002
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Study on Integrated Processing of Speech and Gesture in Multimodal Communication
-
批准号:10480083
-
项目类别:Grant-in-Aid for Scientific Research (B).
-
资助金额:$5.89万
-
财政年份:1998
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Research on application to language education of multimodal ICAI system
-
批准号:07458075
-
项目类别:Grant-in-Aid for Scientific Research (B)
-
资助金额:$3.71万
-
财政年份:1995
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Studies on CAD system of Application Specific VLSI Circuits for Signal Processing
-
批准号:03452174
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$4.22万
-
财政年份:1993
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Co-Operative Study on Modeling and Machine Inplementation of Spoken Language Conversation
-
批准号:02305010
-
项目类别:Grant-in-Aid for Co-operative Research (A)
-
资助金额:$4.67万
-
财政年份:1990
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Application of an Intelligent CAI System with Graphical and Voice Media to Educaton in University for Developmental Scientifical Research
-
批准号:01880035
-
项目类别:Grant-in-Aid for Developmental Scientific Research
-
资助金额:$10.94万
-
财政年份:1989
-
负责人:SHIRAI Katsuhiko
-
依托单位:
Research on the Architecture of the Speech Recognition System and the Computer Aided Design System for Signal Processing LSIs.
-
批准号:61460135
-
项目类别:Grant-in-Aid for General Scientific Research (B)
-
资助金额:$2.5万
-
财政年份:1986
-
负责人:SHIRAI Katsuhiko
-
依托单位:
海外基金