课题基金 / 基金详情

Multilevel Architectures and Algorithms in Deep Learning

Multilevel Architectures and Algorithms in Deep Learning
深度学习中的多级架构和算法
批准号:
464103607
负责人:
Professor Dr. Roland Herzog
金额:
$0.0万
依托单位国家:
德国
项目类别:
Priority Programmes
财政年份:
--
资助国家:
德国
项目状态:
未结题
起止时间:

项目摘要

项目成果

Professor Dr. Roland Herzog的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
The design of deep neural networks (DNNs) and their training is a central issue in machine learning. Progress in these areas is one of the driving forces for the success of these technologies. Nevertheless, tedious experimentation and human interaction is often still needed during the learning process to find an appropriate network structure and corresponding hyperparameters to obtain the desired behavior of a DNN. The strategic goal of the proposed project is to provide algorithmic means to improve this situation. Our methodical approach relies on well established mathematical techniques: identify fundamental algorithmic quantities and construct a-posteriori estimates for them, identify and consistently exploit an appropriate topological framework for the given problem class, establish a multilevel structure for DNNs to account for the fact that DNNs only realize a discrete approximation of a continuous nonlinear mapping relating input to output data. Combining this idea with novel algorithmic control strategies and preconditioning, we will establish the new class of adaptive multilevel algorithms for deep learning, which not only optimize a fixed DNN, but also adaptively refine and extend the DNN architecture during the optimization loop. This concept is not restricted to a particular network architecture, and we will study feedforward neural networks, ResNets, and PINNs as relevant examples. Our integrated approach will thus be able to replace many of the current manual tuning techniques by algorithmic strategies, based on a-posteriori estimates. Moreover, our algorithm will reduce the computational effort for training and also the size of the resulting DNN, compared to a manually designed counterpart, making the use of deep learning more efficient in many aspects. Finally, in the long run our algorithmic approach has the potential to enhance the reliability and interpretability of the resulting trained DNN.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
A Calculus for Non-Smooth Shape Optimization with Applications to Geometric Inverse Problems
Optimal Control of Dissipative Solids: Viscosity Limits and Non-Smooth Algorithms
Impulse Control Problems and Adaptive Numerical Solution of Quasi-Variational Inequalities in Markovian Factor Models
Preconditioned SQP solvers for nonlinear optimization problems with partial differential equations
海外基金