课题基金 / 基金详情

CAREER: Telling the Story of a Visual World: Event Classification and Integrated Image Understanding

CAREER: Telling the Story of a Visual World: Event Classification and Integrated Image Understanding
职业:讲述视觉世界的故事:事件分类和集成图像理解
批准号:
0845230
负责人:
Fei-Fei Li
金额:
$54.88万
依托单位:
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2009
资助国家:
美国
项目状态:
已结题
起止时间:
2009-07-15 至 2010-02-28

项目摘要

项目成果

Fei-Fei Li的其他基金

相似基金

相关文献

中文摘要
翻译
从视觉世界中获得意义的能力,比如识别物体、场景和语义上有意义的活动和事件,是人工智能的基石。在计算机视觉领域,最近在对象级和场景级识别方面取得了非常重要的进展。但是这样的任务通常是在没有完整连贯的场景描述的情况下完成的。此外,目前很少有算法能够进一步解释图像(如事件或活动)的高级语义含义。这个项目的目标是在给定单个未知图像的情况下,通过集成图像理解来实现事件分类。该项目旨在通过开发适合训练算法的复杂学习框架,利用大量来自互联网的真实世界数据,推动对图像的综合和描述性理解的前沿。高精度的性能、最少的人为监督、灵活性和可扩展的学习将是这一努力的重点。这个项目吗?S的理论框架将计算机视觉的几个领域联系在一起,为机器学习领域提供了有趣的模型表示,并将更多的语义驱动的视觉识别问题与自然语言处理领域联系起来。研究结果对视障人士图像理解技术具有重要意义;大型数字图书馆图像的自动标注以及下一代图像检索引擎;以及语言学生和医疗病人(如失语、中风等)的翻译、教育和康复技术。这个项目吗?S的长期教育计划侧重于将最新的视觉计算和认知研究直接带入课堂和整个社区,重点关注未被充分代表的学生群体。
英文摘要
The ability to make meaning out of a visual world, such as recognizing objects, scenes and semantically meaningful activities and events, is a cornerstone of artificial intelligence. In computer vision, very important progress has been made recently in object and scene level recognition. But such tasks are often performed without an integrated and coherent description of the scene. Moreover, very few current algorithms are capable of further interpreting higher level semantic meanings of an image such as an event or activity. The goal of this project is to achieve event classification via an integrated image understanding given a single unknown image.This project aims to push the frontier of integrated and descriptive understanding of images through the development of sophisticated learning frameworks suitable for training algorithms by using a large amount of real-world data such as the ones from the Internet. High accuracy performance, minimal human supervision, flexibility and scalable learning will be the focus of this endeavor. This project?s theoretical framework ties together several areas of computer vision, offers interesting model representations for the machine learning field, and connects more semantically driven visual recognition problem with the natural language processing field. The results are vital for image understanding technology for the visually- impaired; automatic annotation of images for large digital library as well as the next generation of image retrieval engines; and translation, education and rehabilitation technology for language students and medical patients (such as aphasia, stroke, etc.). The project?s long-term educational plan focuses on bringing the latest visual computation and cognition research directly into the classroom and the community at large, with an emphasis on reaching the underrepresented groups of students.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
III: Small: Collaborative Research: Using Large-Scale Image Data for Online Social Media Analysis
  • 批准号:
    1115493
  • 项目类别:
    Standard Grant
  • 资助金额:
    $29.58万
  • 财政年份:
    2011
  • 负责人:
    Fei-Fei Li
  • 依托单位:
CAREER: Telling the Story of a Visual World: Event Classification and Integrated Image Understanding
  • 批准号:
    1000845
  • 项目类别:
    Continuing Grant
  • 资助金额:
    $49.86万
  • 财政年份:
    2009
  • 负责人:
    Fei-Fei Li
  • 依托单位:
Collaborative Research: 1st Sino-USA Summer School in Vision, Learning, Pattern Recognition VLPR 2009
  • 批准号:
    0940687
  • 项目类别:
    Standard Grant
  • 资助金额:
    $2.45万
  • 财政年份:
    2009
  • 负责人:
    Fei-Fei Li
  • 依托单位:
海外基金