课题基金 / 基金详情

CAREER: Theoretical foundations of neural networks - representation, optimization, and generalization

CAREER: Theoretical foundations of neural networks - representation, optimization, and generalization
职业:神经网络的理论基础——表示、优化和泛化
批准号:
1750051
负责人:
Matus Telgarsky
金额:
$50.0万
依托单位国家:
美国
项目类别:
Continuing Grant
财政年份:
2018
资助国家:
美国
项目状态:
已结题
起止时间:
2018-03-15 至 2024-02-29

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Neural networks form the backbone of machine learning's recent advances and sudden ubiquity. Despite this extensive empirical progress, however, a satisfactory understanding of their behavior is still missing. As neural networks enter more and more into human-facing services (self-driving cars, medical diagnostics, etc.), this status quo and in particular its safety ramifications becomes worrisome. This project aims for a theoretical understanding of the foundations of neural networks, divided into three pieces: (a) the representation question regarding which phenomena can be succinctly approximated by neural networks; (b) the optimization question of how to efficiently fit neural networks to data; and (c) the generalization question on why neural networks can fit not only the data they have seen but also the data they have not seen. Developing this understanding will form the core of this project's three broader impacts: (1) the research component will aim to improve safety and reliability of user-facing deployments of neural networks; (2) as an educational component, the research will be simplified and incorporated into freely available course notes; (3) the award supports two outreach efforts co-founded by the PI: UIUC-ML, a university-wide ML seminar; and the midwest ML symposium, a yearly midwest ML gathering.In more detail, the technical focus of this project, divided into the three learning theoretic topics above, is as follows. The core representation question is: what makes neural network representation special? In more detail, the proposed representation questions are firstly to characterize the power gained by adding a single layer to a network, and secondly to characterize the representation properties of recurrent neural networks, namely neural networks which evolve their state along with a time series they consume. Next comes the topic of optimization, where the key mystery is how neural networks manage to perfectly fit their data with simple iterative descent schemes, despite the apparent nonconvexity of the problem. The plan here is to establish an even stronger property: these iterative schemes manage to output networks which not only fit their data, but do so confidently, in the classical sense of margin theory. Finally, the proposal closes with the topic of generalization. The first goal is to develop refined generalization bounds to the point that they can be algorithmically enforced via effective regularization schemes, and secondarily to apply these techniques to the fitting of neural networks to probability distributions, specifically the problem of training Generative Adversarial Networks.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(3)
专著(0)
科研奖励(0)
会议论文
DOI: --
发表时间: 2021-10
期刊: ArXiv
影响因子: --
作者: [Yuzheng Hu;Ziwei Ji;Matus Telgarsky]
通讯作者: Yuzheng Hu;Ziwei Ji;Matus Telgarsky
DOI: --
发表时间: 2021-06
期刊:
影响因子: --
作者: [Ziwei Ji;Justin D. Li;Matus Telgarsky]
通讯作者: Ziwei Ji;Justin D. Li;Matus Telgarsky
DOI: --
发表时间: 2021-07
期刊:
影响因子: --
作者: [Ziwei Ji;N. Srebro;Matus Telgarsky]
通讯作者: Ziwei Ji;N. Srebro;Matus Telgarsky
海外基金