BIGDATA: F: Optimization in Federated Networks of Devices
BIGDATA: F: Optimization in Federated Networks of Devices
批准号:
1838017
负责人:
Ameet Talwalkar
金额:
$99.94万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2019
资助国家:
美国
项目状态:
已结题
起止时间:
2019-01-01 至 2023-12-31
中文摘要
现代远程设备网络(例如移动的电话、可穿戴设备和自动驾驶车辆)每天生成大量数据。这些丰富的数据有可能为各种基于统计机器学习的应用提供动力,例如学习移动的手机用户的活动,适应自动驾驶汽车中的行人行为,预测可穿戴设备中的低血糖等健康事件,或检测智能家居中的故障。由于远程设备的存储和计算能力不断增长,以及与个人数据相关的隐私问题,直接在每个设备上存储和处理数据越来越有吸引力。在新兴的“联邦学习”领域,其目标是使用中央服务器从存储在这些远程设备上的数据中学习统计模型,同时依赖于每个设备的大量计算。联邦学习可以通过数学优化的透镜自然地投射,数学优化是制定和训练大多数机器学习模型的关键组成部分。该项目的重点是解决与联邦优化相关的几个独特的统计和系统挑战。作为该项目的一部分,还开发了一个新的开源基准框架,以具体定义联邦学习中的研究挑战,并促进经验评估的可重复性。这一项目涉及代表性不足人口的学生的参与。 该项目的重点是开发一套新颖的优化方法,以解决在远程设备上学习的独特挑战,包括(a)远程设备和中央服务器之间昂贵的通信;(B)数据、计算资源和设备间通信带宽的高度可变性;以及(c)在任何一个时间只有非常小的一部分远程设备参与培训过程。虽然已经提出了数据中心设置中的许多优化方法来解决(a),但是没有一个在(B)和(c)方面允许显著的灵活性。此外,最近引入的联邦方法数量有限,要么缺乏理论上的收敛保证,要么不能充分解决这三个挑战。该项目旨在开发一套联邦优化方法来解决这些问题,特别是开发和理解凸优化,非凸优化和网络感知优化的技术。这些方法将释放联邦网络的计算能力,以训练高度准确的预测模型,同时遵守严格的系统,网络和隐私限制。该项目利用了优化,统计,机器学习,分布式计算和传感器网络的想法。 除了开发基本的联邦优化方法外,该项目的更广泛影响还包括创建一个新的开源基准框架。该奖项反映了NSF的法定使命,并通过使用基金会的知识价值和更广泛的影响审查标准进行评估,被认为值得支持。
英文摘要
Modern networks of remote devices, such as mobile phones, wearable devices, and autonomous vehicles, generate massive amounts of data each day. This rich data has the potential to power a wide range of statistical machine learning-based applications, such as learning the activities of mobile phone users, adapting to pedestrian behavior in autonomous vehicles, predicting health events like low blood sugar from wearable devices, or detecting burglaries within smart homes. Due to the growing storage and computational power of remote devices, as well as privacy concerns associated with personal data, it is increasingly attractive to store and process data directly on each device. In the burgeoning field of "federated learning," the aim is to use a central server to learn statistical models from data stored across these remote devices, while relying on substantial computation from each device. Federated learning can be naturally cast through the lens of mathematical optimization, a key component in formulating and training most machine learning models. This project focuses on tackling several of the unique statistical and systems challenges associated with federated optimization. As part of this project, a novel open-source benchmarking framework is also being developed to concretely define the research challenges in federated learning and promote reproducibility in empirical evaluations. This project involves participation from students from underrepresented populations. The focus of this project is to develop a novel suite of optimization methods to tackle the unique challenges of learning on remote devices, including (a) expensive communication between remote devices and a central server; (b) high variability in data, computational resources, and communication bandwidth across devices; and (c) a very small fraction of remote devices participating in the training process at any one time. While numerous optimization methods in the data center setting have been proposed to tackle (a), none allow significant flexibility in terms of (b) and (c). Further, the limited number of recently introduced federated methods either lack theoretical convergence guarantees or do not adequately address these three challenges. This project aims to develop a suite of federated optimization methods to tackle these issues, specifically developing and understanding techniques for: convex optimization, non-convex optimization, and network-aware optimization. These methods will unleash the computational power of federated networks to train highly-accurate predictive models while adhering to strict systems, network, and privacy constraints. This project leverages ideas from optimization, statistics, machine learning, distributed computing, and sensor networks. In addition to developing foundational federated optimization methods, the broader impact of this project includes the creation of a novel open-source benchmarking framework.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(11)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
--
发表时间:
2019-06
期刊:
影响因子:
--
作者:
[M. Khodak;Maria-Florina Balcan;Ameet Talwalkar]
通讯作者:
M. Khodak;Maria-Florina Balcan;Ameet Talwalkar
DOI:
--
发表时间:
2019-02
期刊:
ArXiv
影响因子:
--
作者:
[M. Khodak;Maria-Florina Balcan;Ameet Talwalkar]
通讯作者:
M. Khodak;Maria-Florina Balcan;Ameet Talwalkar
DOI:
--
发表时间:
2020-12
期刊:
影响因子:
--
作者:
[Tian Li;Shengyuan Hu;Ahmad Beirami;Virginia Smith]
通讯作者:
Tian Li;Shengyuan Hu;Ahmad Beirami;Virginia Smith
DOI:
--
发表时间:
2022
期刊:
影响因子:
--
作者:
[Ravikumar Balakrishnan;Tian Li;Tianyi Zhou;N. Himayat;Virginia Smith;J. Bilmes]
通讯作者:
Ravikumar Balakrishnan;Tian Li;Tianyi Zhou;N. Himayat;Virginia Smith;J. Bilmes
DOI:
--
发表时间:
2021-06
期刊:
ArXiv
影响因子:
--
作者:
[Zachary B. Charles;Zachary Garrett;Zhouyuan Huo;Sergei Shmulyian;Virginia Smith]
通讯作者:
Zachary B. Charles;Zachary Garrett;Zhouyuan Huo;Sergei Shmulyian;Virginia Smith
共 11 条
Travel: NSF Student Travel Grant for the Sixth Conference on Machine Learning and Systems (MLSys 2023)
-
批准号:2325547
-
项目类别:Standard Grant
-
资助金额:$5.0万
-
财政年份:2023
-
负责人:Ameet Talwalkar
-
依托单位:
CAREER: Foundations of Next-Generation Neural Architecture Search
-
批准号:2046613
-
项目类别:Continuing Grant
-
资助金额:$55.0万
-
财政年份:2021
-
负责人:Ameet Talwalkar
-
依托单位:
Model-Parallel Collaborative Filtering in Apache Spark
-
批准号:1555772
-
项目类别:Standard Grant
-
资助金额:$6.88万
-
财政年份:2015
-
负责人:Ameet Talwalkar
-
依托单位:
SIFTER: A Systems Biology Platform for Protein Function Prediction
-
批准号:1122732
-
项目类别:Fellowship Award
-
资助金额:$24.0万
-
财政年份:2011
-
负责人:Ameet Talwalkar
-
依托单位:
国内基金
海外基金
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis
-
批准号:--
-
项目类别:合作创新研究团队
-
资助金额:--
-
批准年份:2024
-
负责人:姚韬
-
依托单位:
供应链管理中的稳健型(Robust)策略分析和稳健型优化(Robust Optimization )方法研究
-
批准号:70601028
-
项目类别:青年科学基金项目
-
资助金额:7.0万元
-
批准年份:2006
-
负责人:王明征
-
依托单位: