课题基金 / 基金详情

Seebibyte: Visual Search for the Era of Big Data

Seebibyte: Visual Search for the Era of Big Data
Seebibyte:大数据时代的视觉搜索
批准号:
EP/M013774/1
负责人:
Andrew Zisserman
金额:
$569.27万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2015
资助国家:
英国
项目状态:
已结题
起止时间:
2015 至 --

项目摘要

项目成果

Andrew Zisserman的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
The Programme is organised into two themes. Research theme one will develop new computer vision algorithms to enable efficient search and description of vast image and video datasets - for example of the entire video archive of the BBC. Our vision is that anything visual should be searchable for, in the manner of a Google search of the web: by specifying a query, and having results returned immediately, irrespective of the size of the data. Such enabling capabilities will have widespread application both for general image/video search - consider how Google's web search has opened up new areas - and also for designing customized solutions for searching.A second aspect of theme 1 is to automatically extract detailed descriptions of the visual content. The aim here is to achieve human like performance and beyond, for example in recognizing configurations of parts and spatial layout, counting and delineating objects, or recognizing human actions and inter-actions in videos, significantly superseding the current limitations of computer vision systems, and enabling new and far reaching applications. The new algorithms will learn automatically, building on recent breakthroughs in large scale discriminative and deep machine learning. They will be capable of weakly-supervised learning, for example from images and videos downloaded from the internet, and require very little human supervision.The second theme addresses transfer and translation. This also has two aspects. The first is to apply the new computer vision methodologies to `non-natural' sensors and devices, such as ultrasound imaging and X-ray, which have different characteristics (noise, dimension, invariances) to the standard RGB channels of data captured by `natural' cameras (iphones, TV cameras). The second aspect of this theme is to seek impact in a variety of other disciplines and industry which today greatly under-utilise the power of the latest computer vision ideas. We will target these disciplines to enable them to leapfrog the divide between what they use (or do not use) today which is dominated by manual review and highly interactive analysis frame-by-frame, to a new era where automated efficient sorting, detection and mensuration of very large datasets becomes the norm. In short, our goal is to ensure that the newly developed methods are used by academic researchers in other areas, and turned into products for societal and economic benefit. To this end open source software, datasets, and demonstrators will be disseminated on the project website.The ubiquity of digital imaging means that every UK citizen may potentially benefit from the Programme research in different ways. One example is an enhanced iplayer that can search for where particular characters appear in a programme, or intelligently fast forward to the next `hugging' sequence. A second is wider deployment of lower cost imaging solutions in healthcare delivery. A third, also motivated by healthcare, is through the employment of new machine learning methods for validating targets for drug discovery based on microscopy images
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
Deep Audio-Visual Speech Recognition
深度视听语音识别
DOI: 10.48550/arxiv.1809.02108
发表时间: 2018
期刊:
影响因子: --
作者: [Afouras T]
通讯作者: Afouras T
DOI: 10.1109/cvpr.2017.313
发表时间: 2016-11
期刊: 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
影响因子: --
作者: [Thalaiyasingam Ajanthan;Alban Desmaison;Rudy Bunel;M. Salzmann;Philip H. S. Torr;M. P. Kumar]
通讯作者: Thalaiyasingam Ajanthan;Alban Desmaison;Rudy Bunel;M. Salzmann;Philip H. S. Torr;M. P. Kumar
DOI: 10.21437/interspeech.2018-1400
发表时间: 2018-04
期刊: ArXiv
影响因子: --
作者: [Triantafyllos Afouras;Joon Son Chung;Andrew Zisserman]
通讯作者: Triantafyllos Afouras;Joon Son Chung;Andrew Zisserman
Now You're Speaking My Language: Visual Language Identification
现在你正在说我的语言:视觉语言识别
DOI: 10.21437/interspeech.2020-2921
发表时间: 2020
期刊:
影响因子: --
作者: [Afouras T]
通讯作者: Afouras T
7
    Visual AI: An Open World Interpretable Visual Transformer
    • 批准号:
      EP/T028572/1
    • 项目类别:
      Research Grant
    • 资助金额:
      $753.32万
    • 财政年份:
      2020
    • 负责人:
      Andrew Zisserman
    • 依托单位:
    Learning to Recognise Dynamic Visual Content from Broadcast Footage
    • 批准号:
      EP/I012001/1
    • 项目类别:
      Research Grant
    • 资助金额:
      $63.82万
    • 财政年份:
      2011
    • 负责人:
      Andrew Zisserman
    • 依托单位:
    国内基金
    海外基金
    基于多幅图象的Visual Hull重构及表面属性建模算法研究
    • 批准号:
      60373031
    • 项目类别:
      面上项目
    • 资助金额:
      23.0万元
    • 批准年份:
      2003
    • 负责人:
      陈越
    • 依托单位: