课题基金 / 基金详情

Systemization of audio-visual knowledge resources using graphical models

Systemization of audio-visual knowledge resources using graphical models
利用图模型将视听知识资源系统化
批准号:
17300059
负责人:
SHINODA Koichi
金额:
$9.46万
依托单位国家:
日本
项目类别:
Grant-in-Aid for Scientific Research (B)
财政年份:
2005
资助国家:
日本
项目状态:
已结题
起止时间:
2005 至 2007

项目摘要

项目成果

SHINODA Koichi的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Recent advances in computer technology, particularly in storage technology, have resulted in significant increases in the number and quality of audio-visual knowledge resources. Most of those resources are not equipped with index information, and thus, it has become difficult for ordinary people to browse the entire content of each database. Techniques for systemizing audio-visual knowledge resources and utilizing them have been strongly demanded. However, statistical pattern recognition techniques have not yet achieved enough performance for this purpose. In addition, it is not always clear what kinds of indexing are useful. In this study, we take an approach to index those databases in different ways with unsupervised manner, and extract dependencies among those labels. First, we carried scene recognition for baseball video. We constructed annotated database for 43 games of Major League Baseball with NHK Science & Technical Research Labs and used them for our evaluation. We used vari … More ous relationships between scene labels such as scene contexts, and unified audio and visual information. We achieved 60% accuracy for 16 scene recognition and 90% recall rate for score scene detection. Our techniques are expected to contribute much to make automatic highlight extraction systems for broadcast companies. Second, we participated in TRECVID workshop organized by NIST, USA, to study the high-level feature extraction task. We constructed tree-structured dictionaries of "visual words" by unsupervised clustering for video features, and selected a tree-cut as a dictionary for each word. By using Bag-of-word approach, we constructed a robust extraction system against the differences in data amount for each feature. We also extracted effective "motion words" for dynamic features. Our method achieved significant improvements in the task of extracting 39 features. The other research topics include robust speech recognition using graphical models, multi-modal interface for asynchronous multi-modal inputs, human-gait modeling. Less
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
「音声とペンの同時入力における個人差への適応化」
“适应同时语音和笔输入的个体差异”
DOI: --
发表时间: 2008
期刊:
影响因子: --
作者: [渡邉 康司, 篠田 浩一, 古井 貞煕]
通讯作者: 古井 貞煕
Model adaptation for semi-synchronous speech and pen input
半同步语音和笔输入的模型自适应
DOI: --
发表时间: 2008
期刊:
影响因子: --
作者: [Yasushi, Watanabe, Koichi, Shinoda, Sadaoki, Furui]
通讯作者: Furui
"TokyoTech's TRECVID2007 Notebook"
“TokyoTech 的 TRECVID2007 笔记本”
DOI: --
发表时间: 2007
期刊:
影响因子: --
作者: [T. Nakamura, K. Shinoda and S. Furui]
通讯作者: K. Shinoda and S. Furui
TokyoTech's TRECVID2007 Notebook
东京工业大学的 TRECVID2007 笔记本
DOI: --
发表时间: 2007
期刊:
影响因子: --
作者: [T., Nakamura, K., Shinoda, S., Furui]
通讯作者: Furui
59
    A study of multimodal recognition for human communication search
    • 批准号:
      20300063
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $11.48万
    • 财政年份:
      2008
    • 负责人:
      SHINODA Koichi
    • 依托单位:
    SPEECH RECOGNITION WITH SYNCHRONOUS INPUT OF HAND-WRITTEN GESTURES FOR MOBILE DEVICES
    • 批准号:
      15300054
    • 项目类别:
      Grant-in-Aid for Scientific Research (B)
    • 资助金额:
      $3.78万
    • 财政年份:
      2003
    • 负责人:
      SHINODA Koichi
    • 依托单位:
    海外基金