Asymptotic analysis of online training algorithms in deep learning
Asymptotic analysis of online training algorithms in deep learning
批准号:
2879209
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2023
资助国家:
英国
项目状态:
未结题
起止时间:
2023 至 --
中文摘要
点击翻译按钮获取中文摘要
英文摘要
Neural networks have achieved immense practical success in various fields of science, engineering and finance due to their ability to learn high-dimensional, nonlinear relationships from large datasets. The parameters of the neural network are calibrated by 'training' the neural network to minimise an appropriate objective function for a dataset using stochastic gradient descent (SGD) methods. Current mathematical understanding of why SGD methods can successfully train a neural network and how such a neural network generalizes to new out-of-sample data is limited to cases where the neural network has a simple architecture. The objectives of the planned research is twofold. First, we will develop mathematical theory for more sophisticated neural network architectures. A notable example is the family of recurrent neural networks (RNNs), which include a hidden state with the "memory" of the data sequence at previous time steps. Novel approaches are required to analyse RNNs. The memory updates depend strongly on the distribution of the input data sequence, which will be correlated across time. Another example is deep reinforcement learning algorithms such as actor-critic neural network algorithms. These algorithms are challenging to mathematically analyse since they are simultaneously learning the dynamics as well as an optimal policy. Furthermore, the distribution of the data changes as the reinforcement learning model changes during training. In our analysis, we plan to study both single-layer and multi-layer (deep) neural networks. Secondly, our analysis will attempt to study important fundamental questions for the implementation of deep learning models in applications, including how information propagates (the vanishing/exploding gradient problem). Our research will contribute to the mathematical theory of deep learning. Convergence and generalization theory for deep learning models is important to guarantee the reliability and accuracy of deep learning when implemented in applications. This project falls within the following EPSRC research areas: non-linear systems, statistics and applied probability, numerical analysis, and mathematical sciences.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
国内基金
海外基金
登录
查看更多内容
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis
-
批准号:--
-
项目类别:合作创新研究团队
-
资助金额:--
-
批准年份:2024
-
负责人:姚韬
-
依托单位:
Intelligent Patent Analysis for Optimized Technology Stack Selection:Blockchain BusinessRegistry Case Demonstration
-
批准号:--
-
项目类别:外国学者研究基金项目
-
资助金额:--
-
批准年份:2024
-
负责人:USHARANI HAREESH GOVINDARA JAN
-
依托单位:
利用全基因组关联分析和QTL-seq发掘花生白绢病抗性分子标记
-
批准号:31971981
-
项目类别:面上项目
-
资助金额:58.0万元
-
批准年份:2019
-
负责人:晏立英
-
依托单位:
基于SERS纳米标签和光子晶体的单细胞Western Blot定量分析技术研究
-
批准号:31900571
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2019
-
负责人:刘兵
-
依托单位:
利用多个实验群体解析猪保幼带形成及其自然消褪的遗传机制
-
批准号:31972542
-
项目类别:面上项目
-
资助金额:57.0万元
-
批准年份:2019
-
负责人:郭源梅
-
依托单位:
基于Meta-analysis的新疆棉花灌水增产模型研究
-
批准号:41601604
-
项目类别:青年科学基金项目
-
资助金额:22.0万元
-
批准年份:2016
-
负责人:赵爱琴
-
依托单位:
基于个体分析的投影式非线性非负张量分解在高维非结构化数据模式分析中的研究
-
批准号:61502059
-
项目类别:青年科学基金项目
-
资助金额:19.0万元
-
批准年份:2015
-
负责人:刘昶
-
依托单位:
多目标诉求下我国交通节能减排市场导向的政策组合选择研究
-
批准号:71473155
-
项目类别:面上项目
-
资助金额:60.0万元
-
批准年份:2014
-
负责人:柴建
-
依托单位:
大规模微阵列数据组的meta-analysis方法研究
-
批准号:31100958
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2011
-
负责人:赵洪雅
-
依托单位:
基于物质流分析的中国石油资源流动过程及碳效应研究
-
批准号:41101116
-
项目类别:青年科学基金项目
-
资助金额:23.0万元
-
批准年份:2011
-
负责人:刘晓洁
-
依托单位: