课题基金 / 基金详情

CAREER: Exploiting Deep Generative Models for Visual Recognition

CAREER: Exploiting Deep Generative Models for Visual Recognition
职业:利用深度生成模型进行视觉识别
批准号:
2239076
负责人:
Jun-Yan Zhu
金额:
$58.19万
依托单位:
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2023
资助国家:
美国
项目状态:
未结题
起止时间:
2023-04-01 至 2028-03-31

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Modern visual recognition systems have achieved impressive results on standard benchmarks and work reliably for common objects and scenes, given massive data and annotations. Unfortunately, current systems struggle to detect rare or unseen objects and fail to adapt to new domains. Researchers, engineers and/or domain experts have to capture and annotate huge amounts of real data, which are costly for common objects and impractical for rare objects and corner cases (i.e., cases that occur when multiple unique conditions simultaneously occur). To address the above challenges and automatically create and label data that fully depict the corner cases, this project leverages the rich compositional structure and powerful synthesis capacity of large-scale generative models. By using these models that can quickly synthesize diverse objects and scenes with an unknown visual elements (e.g., new poses, weather, lighting, etc.). This project will develop recognition algorithms that can recognize rare/unseen objects to adapt to continuously changing environments. This project has a potential to be transformative for various applications, such as autonomous driving, assistive robots, healthcare, e-commerce, and mixed reality. Furthermore, this research will translate to code, models, courses, and tutorials, that are widely accessible to diverse stakeholders and education and research programs that engage with the broader community. Directly using generative models is challenging, as it is highly unlikely that a randomly sampled image will cover a corner case that can improve recognition systems. To synthesize data that more closely resemble the long-tail distribution and new domains, this project will focus on three research thrusts. First, the project addresses learning visual recognition via generative models by exploring different methods of automatically generating data and annotations. Second, the project will analyze visual recognition systems through generative models by synthesizing diverse, continuously evolving test data to interrogate the system and understand the biases. Finally, the project will automatically select and adapt generative models to new domains and tasks. These three thrusts are tightly connected, as once the algorithms identify hard examples that fail our current system, these examples can be used to close the loop between training and analysis. Finally, investigators will evaluate the developed method by comparing methods with or without using generative models.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(3)
专著(0)
科研奖励(0)
会议论文
DOI: 10.1109/iccv51070.2023.02074
发表时间: 2023-03
期刊: 2023 IEEE/CVF International Conference on Computer Vision (ICCV)
影响因子: --
作者: [Nupur Kumari;Bin Zhang;Sheng-Yu Wang;Eli Shechtman;Richard Zhang;Jun-Yan Zhu]
通讯作者: Nupur Kumari;Bin Zhang;Sheng-Yu Wang;Eli Shechtman;Richard Zhang;Jun-Yan Zhu
DOI: 10.1109/iccv51070.2023.00694
发表时间: 2023-04
期刊: 2023 IEEE/CVF International Conference on Computer Vision (ICCV)
影响因子: --
作者: [Songwei Ge;Taesung Park;Jun-Yan Zhu;Jia-Bin Huang]
通讯作者: Songwei Ge;Taesung Park;Jun-Yan Zhu;Jia-Bin Huang
DOI: 10.1145/3610548.3618189
发表时间: 2022-10
期刊: SIGGRAPH Asia 2023 Conference Papers
影响因子: --
作者: [Daohan Lu;Sheng-Yu Wang;Nupur Kumari;Rohan Agarwal;David Bau;Jun-Yan Zhu]
通讯作者: Daohan Lu;Sheng-Yu Wang;Nupur Kumari;Rohan Agarwal;David Bau;Jun-Yan Zhu
海外基金