课题基金 / 基金详情

Asymptotic analysis of online training algorithms in deep learning

Asymptotic analysis of online training algorithms in deep learning
深度学习在线训练算法的渐近分析
批准号:
2879209
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2023
资助国家:
英国
项目状态:
未结题
起止时间:
2023 至 --

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Neural networks have achieved immense practical success in various fields of science, engineering and finance due to their ability to learn high-dimensional, nonlinear relationships from large datasets. The parameters of the neural network are calibrated by 'training' the neural network to minimise an appropriate objective function for a dataset using stochastic gradient descent (SGD) methods. Current mathematical understanding of why SGD methods can successfully train a neural network and how such a neural network generalizes to new out-of-sample data is limited to cases where the neural network has a simple architecture. The objectives of the planned research is twofold. First, we will develop mathematical theory for more sophisticated neural network architectures. A notable example is the family of recurrent neural networks (RNNs), which include a hidden state with the "memory" of the data sequence at previous time steps. Novel approaches are required to analyse RNNs. The memory updates depend strongly on the distribution of the input data sequence, which will be correlated across time. Another example is deep reinforcement learning algorithms such as actor-critic neural network algorithms. These algorithms are challenging to mathematically analyse since they are simultaneously learning the dynamics as well as an optimal policy. Furthermore, the distribution of the data changes as the reinforcement learning model changes during training. In our analysis, we plan to study both single-layer and multi-layer (deep) neural networks. Secondly, our analysis will attempt to study important fundamental questions for the implementation of deep learning models in applications, including how information propagates (the vanishing/exploding gradient problem). Our research will contribute to the mathematical theory of deep learning. Convergence and generalization theory for deep learning models is important to guarantee the reliability and accuracy of deep learning when implemented in applications. This project falls within the following EPSRC research areas: non-linear systems, statistics and applied probability, numerical analysis, and mathematical sciences.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
国内基金
海外基金
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis
Intelligent Patent Analysis for Optimized Technology Stack Selection:Blockchain BusinessRegistry Case Demonstration
  • 批准号:
    --
  • 项目类别:
    外国学者研究基金项目
  • 资助金额:
    --
  • 批准年份:
    2024
  • 负责人:
    USHARANI HAREESH GOVINDARA JAN
  • 依托单位:
利用全基因组关联分析和QTL-seq发掘花生白绢病抗性分子标记
基于SERS纳米标签和光子晶体的单细胞Western Blot定量分析技术研究
  • 批准号:
    31900571
  • 项目类别:
    青年科学基金项目
  • 资助金额:
    24.0万元
  • 批准年份:
    2019
  • 负责人:
    刘兵
  • 依托单位: