Collaborative Research: RI:Medium:Understanding Events from Streaming Video - Joint Deep and Graph Representations, Commonsense Priors, and Predictive Learning
Collaborative Research: RI:Medium:Understanding Events from Streaming Video - Joint Deep and Graph Representations, Commonsense Priors, and Predictive Learning
批准号:
1956050
负责人:
Sudeep Sarkar
金额:
$42.12万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2020
资助国家:
美国
项目状态:
已结题
起止时间:
2020-10-01 至 2024-09-30
中文摘要
点击翻译按钮获取中文摘要
英文摘要
While it is easy for humans to process video data and extract meanings from it, it is extremely hard to design algorithms to do so. When developed, there are many applications of this technology, such as building assistive robotics or constructing smart spaces for independent living or monitoring wildlife. Video-data capture events, which are central to the content of human experience. Events consist of objects/people (who), location (where), time (when), actions (what), activities (how), and intent (why). This project develops a computer vision-based event understanding algorithm that operates in a self-supervised, streaming fashion. The algorithm will predict and detect old and new events, learn to build hierarchical event representations, all in the context of a prior knowledge-base that is updated over time. The intent is to generate interpretations of an event that go beyond what is seen, rather than just recognition. This research pushes the frontier of computer vision by coupling the self-supervised learning process with prior knowledge, moving the field towards open-world algorithms, and needing little or no supervision. Furthermore, this project will focus on recruitment and retention of undergraduate women students through freshman and sophomore years, with attention towards underrepresented minority students at the three sites: University of South Florida, Florida State University, and Oklahoma State University.At the core of the approach is a hybrid representational hierarchy that includes both continuous representations and symbolic graph-based representations. The continuous-valued representation is the standard, vector-valued deep learning stack that ends in an embedding vector of some object or action concept in the knowledge base. The next level of the representation consists of elementary symbolic compositions of these verbs and nouns. These elementary compositions, when associated with concepts from a knowledge-base they makeup an event interpretation, containing descriptions that go beyond what is observed in the image. These symbolic levels are built using Grenander's canonical representations from pattern theory. These representations, which have flexible graph-structured backbones, are more expressive than other well-known graphical models. The specific technical aims of the project are four-fold. First, it seeks to integrate function-based continuous with energy-based Grenander's canonical symbolic representations from pattern theory into one integrated formulation based on equilibrium propagation. Second, it will research and develop ways to use and modify commonsense knowledge bases. This will help to go beyond the closed world assumption, which is implicit in the current practice of annotated data-based deep learning approaches. Third, it will develop dynamical models on graph manifolds, which will enable generative modeling of graph structures for prediction and discovery of new concepts. Fourth, inspired by finding from human perception experiments and neuroscience, it will design predictive self-supervised learning over both continuous and symbolic representations.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(8)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
Spatio-Temporal Event Segmentation for Wildlife Extended Videos
野生动物扩展视频的时空事件分割
DOI:
10.1007/978-3-031-11349-9
发表时间:
2022
期刊:
International Conference on Computer Vision and Image Processing
影响因子:
--
作者:
[Mounir, R., Gula, R., J., Sarkar]
通讯作者:
J., Sarkar
DOI:
10.1007/s10851-021-01027-1
发表时间:
2021-03
期刊:
Journal of Mathematical Imaging and Vision
影响因子:
2
作者:
[Xiaoyang Guo;A. Srivastava;S. Sarkar]
通讯作者:
Xiaoyang Guo;A. Srivastava;S. Sarkar
DOI:
10.1007/978-3-031-19839-7_5
发表时间:
2021-04
期刊:
影响因子:
--
作者:
[Sathyanarayanan N. Aakur;Sudeep Sarkar]
通讯作者:
Sathyanarayanan N. Aakur;Sudeep Sarkar
DOI:
10.1007/978-3-031-19833-5_26
发表时间:
2022
期刊:
影响因子:
--
作者:
[A. Bal;R. Mounir;Sathyanarayanan N. Aakur;Sudeep Sarkar;Anuj Srivastava]
通讯作者:
A. Bal;R. Mounir;Sathyanarayanan N. Aakur;Sudeep Sarkar;Anuj Srivastava
DOI:
10.1109/tpami.2023.3287837
发表时间:
2023-06
期刊:
IEEE Transactions on Pattern Analysis and Machine Intelligence
影响因子:
23.6
作者:
[Sathyanarayanan N. Aakur;Sudeep Sarkar]
通讯作者:
Sathyanarayanan N. Aakur;Sudeep Sarkar
共 7 条
I-Corps Sites: Type II - I-Corps Site at University of South Florida Tampa
-
批准号:1829217
-
项目类别:Continuing Grant
-
资助金额:$16.0万
-
财政年份:2018
-
负责人:Sudeep Sarkar
-
依托单位:
I-Corps: Semantic Video - from Video to Descriptions
-
批准号:1647887
-
项目类别:Standard Grant
-
资助金额:$5.0万
-
财政年份:2016
-
负责人:Sudeep Sarkar
-
依托单位:
I-Corps Sites: University of South Florida: Catalyzing Research Translation
-
批准号:1449137
-
项目类别:Continuing Grant
-
资助金额:$29.97万
-
财政年份:2015
-
负责人:Sudeep Sarkar
-
依托单位:
RI: Small: Collaborative Research: Ontology based Perceptual Organization of Audio-Video Events using Pattern Theory
-
批准号:1217676
-
项目类别:Standard Grant
-
资助金额:$24.98万
-
财政年份:2012
-
负责人:Sudeep Sarkar
-
依托单位:
EMT/Nano: Energy Minimization Computing using Field Coupled Nanomagnets--Modeling and Fabrication
-
批准号:0829838
-
项目类别:Standard Grant
-
资助金额:$25.0万
-
财政年份:2008
-
负责人:Sudeep Sarkar
-
依托单位:
ITR: Fundamental Issues in Automated American Sign Language Recognition
-
批准号:0312993
-
项目类别:Continuing Grant
-
资助金额:$37.98万
-
财政年份:2003
-
负责人:Sudeep Sarkar
-
依托单位:
CISE Research Resources: A Compute-Intensive Sensor-Based Environment for Research in Computer Vision and Artificial Intelligence
-
批准号:0130768
-
项目类别:Standard Grant
-
资助金额:$14.12万
-
财政年份:2001
-
负责人:Sudeep Sarkar
-
依托单位:
Enhancing Undergraduate Computer Science Curriculum through Image Computations: Proof-of-Concept
-
批准号:9980832
-
项目类别:Standard Grant
-
资助金额:$7.52万
-
财政年份:2000
-
负责人:Sudeep Sarkar
-
依托单位:
The Role Learning in Perceptual Organization of Complex Images
-
批准号:9907141
-
项目类别:Continuing Grant
-
资助金额:$22.57万
-
财政年份:1999
-
负责人:Sudeep Sarkar
-
依托单位:
Major Research Instrumentation: Acquisition of a Cyberware 3D Scanner to Facilitate State of Art Research in Computer Vision and Graphics
-
批准号:9724422
-
项目类别:Standard Grant
-
资助金额:$11.5万
-
财政年份:1997
-
负责人:Sudeep Sarkar
-
依托单位:
CAREER: The Role of Perceptual Organization in Motion Analysis
-
批准号:9501932
-
项目类别:Continuing Grant
-
资助金额:$14.49万
-
财政年份:1995
-
负责人:Sudeep Sarkar
-
依托单位:
国内基金
海外基金
登录
查看更多内容
Research on Quantum Field Theory without a Lagrangian Description
-
批准号:24ZR1403900
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:SATOSHI NAWATA
-
依托单位:
Cell Research
-
批准号:31224802
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2012
-
负责人:程磊
-
依托单位:
Cell Research
-
批准号:31024804
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2010
-
负责人:程磊
-
依托单位:
Cell Research (细胞研究)
-
批准号:30824808
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2008
-
负责人:张爱兰
-
依托单位:
Research on the Rapid Growth Mechanism of KDP Crystal
-
批准号:10774081
-
项目类别:面上项目
-
资助金额:45.0万元
-
批准年份:2007
-
负责人:滕冰
-
依托单位: