CPS: Small: Distributed Learning for Control of Cyber-Physical Systems
CPS: Small: Distributed Learning for Control of Cyber-Physical Systems
批准号:
1932011
负责人:
Michael Zavlanos
金额:
$40.75万
依托单位:
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2019
资助国家:
美国
项目状态:
已结题
起止时间:
2019-10-01 至 2023-09-30
中文摘要
点击翻译按钮获取中文摘要
英文摘要
In state-of-the-art Cyber-Physical-Systems (CPS) supervised learning or unsupervised learning are typically used to analyze data. Nevertheless, in many such systems rules cannot be determined in advance and these data mining techniques are not directly applicable due to the dynamic nature of the data, their large volume that prohibits labelling in practice, and the fact that these data are added to the system piece by piece and not altogether in advance. On the other hand, control of CPS is usually done in a model-based manner, where a desired control policy is computed from a high-fidelity system model that has been derived at design-time, and potentially may be updated at runtime. However, this approach is not suitable for highly dynamical CPS, that potentially represent systems of systems whose spatial and temporal configurations may rapidly change. In fact, with such high number of configuration levels, it is almost impossible to derive suitable control policies using standard model-driven techniques. Consequently, it is critical to facilitate design of data-based controllers, with strong performance guarantees, in a way that allows for natural runtime control adaptation. Reinforcement Learning (RL) provides such a framework. In RL agents interact with the environment in a feedback loop to learn an optimal policy by taking appropriate sequences of actions in order to optimize longterm payoff. As such, RL can be much more efficient compared to supervised and unsupervised learning, in analyzing streaming data and especially in controlling a system. The goal of this project is to develop a distributed off-policy RL framework for the control of CPS. Distributed RL methods avoid the fragility, communication overhead, and privacy concerns of collecting all information at a central processing unit. Moreover, off-policy learning methods significantly improve sampling efficiency and ensure safer operation. The distributed RL framework developed under this project will have a profound impact on the control of CPS, in areas as diverse as transportation, manufacturing, health-care, smart city, urban planning, etc., that rely on multiple sensors for data collection and control. This project also involves an educational agenda focusing on K-12, undergraduate, and graduate level education. The outreach component of this project focuses on improving the pre-college students' awareness of the potential and attractiveness of a research and engineering career.The technical aims of this project are divided into four thrusts. The first thrust develops distributed off-policy RL methods using linear function approximation of the action-value function. Distributed RL algorithms using linear function approximation have been proposed for policy evaluation only. This thrust develops new RL algorithms that can also improve the policy until an optimal policy is found, which is necessary for control. Since defining appropriate feature vectors for RL problems is generally difficult and since linear mappings might not able to capture possibly nonlinear interactions between these features, the second thrust develops distributed off-policy RL methods using nonlinear function approximation, specifically, Neural Networks. The third thrust develops distributed off-policy Actor-Critic methods. When the action space is large or continuous, Actor-Critic methods are much more effective since they parameterize the target policy function using either linear or nonlinear function approximation and learn the optimal parameter so that the resulting policy maps to the optimal action for every state. Finally, the fourth thrust develops distributed RL methods for asynchronous, heterogeneous, and non-stationary data that are common in modern CPS, where sensors do not observe identically distributed data nor do they sample data at the same time. Moreover, the distributions from which data are sampled can change with time. This project focuses on the development of algorithms and supporting theoretical results. The developed algorithms are evaluated in simulation on resource allocation problems in CPS, specifically, on the control of distributed shared vehicle dispatch systems.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(18)
专著(0)
科研奖励(0)
会议论文
登录
查看更多内容
DOI:
10.1016/j.automatica.2020.109218
发表时间:
2020
期刊:
Automatica
影响因子:
6.4
作者:
[Zhang, Yan, Zavlanos, Michael M.]
通讯作者:
Zavlanos, Michael M.
DOI:
10.48550/arxiv.2203.08957
发表时间:
2022-03
期刊:
ArXiv
影响因子:
--
作者:
[Zifan Wang;Yi Shen;M. Zavlanos]
通讯作者:
Zifan Wang;Yi Shen;M. Zavlanos
DOI:
10.1109/tro.2020.2980176
发表时间:
2018-12
期刊:
IEEE Transactions on Robotics
影响因子:
7.8
作者:
[Reza Khodayi-mehr;M. Zavlanos]
通讯作者:
Reza Khodayi-mehr;M. Zavlanos
Transfer Reinforcement Learning under Unobserved Contextual Information
未观察到的上下文信息下的迁移强化学习
DOI:
10.1109/iccps48487.2020.00015
发表时间:
2020
期刊:
2020 ACM/IEEE 11th International Conference on Cyber-Physical Systems (ICCPS
影响因子:
--
作者:
[Zhang, Yan, Zavlanos, Michael M.]
通讯作者:
Zavlanos, Michael M.
DOI:
--
发表时间:
2021-06
期刊:
ArXiv
影响因子:
--
作者:
[Panagiotis Vlantis;M. Zavlanos]
通讯作者:
Panagiotis Vlantis;M. Zavlanos
共 16 条
CPS: Medium: Collaborative Research: Human-on-the-Loop Control for Smart Ultrasound Imaging
-
批准号:1837499
-
项目类别:Standard Grant
-
资助金额:$60.0万
-
财政年份:2018
-
负责人:Michael Zavlanos
-
依托单位:
NeTS: Medium: Collaborative Research: Optimal Communication for Faster Sensor Network Coordination
-
批准号:1302284
-
项目类别:Standard Grant
-
资助金额:$26.0万
-
财政年份:2013
-
负责人:Michael Zavlanos
-
依托单位:
NeTS: Synergy: Collaborative Research: Controlling Teams of Autonomous Mobile Beamformers
-
批准号:1239339
-
项目类别:Standard Grant
-
资助金额:$29.4万
-
财政年份:2013
-
负责人:Michael Zavlanos
-
依托单位:
RI: Medium: Collaborative Research: Mobile Microrobot Platform for Advanced Manufacturing Applications
-
批准号:1302283
-
项目类别:Continuing Grant
-
资助金额:$18.45万
-
财政年份:2013
-
负责人:Michael Zavlanos
-
依托单位:
CAREER: Control of Mobile Robot Networks: Integrating the Communication and Physical Domains
-
批准号:1261828
-
项目类别:Continuing Grant
-
资助金额:$34.58万
-
财政年份:2012
-
负责人:Michael Zavlanos
-
依托单位:
CAREER: Control of Mobile Robot Networks: Integrating the Communication and Physical Domains
-
批准号:1054604
-
项目类别:Continuing Grant
-
资助金额:$44.96万
-
财政年份:2011
-
负责人:Michael Zavlanos
-
依托单位:
国内基金
海外基金
登录
查看更多内容
昼夜节律性small RNA在血斑形成时间推断中的法医学应用研究
-
批准号:
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:
-
依托单位:
tRNA-derived small RNA上调YBX1/CCL5通路参与硼替佐米诱导慢性疼痛的机制研究
-
批准号:
-
项目类别:省市级项目
-
资助金额:10.0万元
-
批准年份:2022
-
负责人:张祥忠
-
依托单位:
Small RNA调控I-F型CRISPR-Cas适应性免疫性的应答及分子机制
-
批准号:32000033
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2020
-
负责人:林平
-
依托单位:
Small RNAs调控解淀粉芽胞杆菌FZB42生防功能的机制研究
-
批准号:31972324
-
项目类别:面上项目
-
资助金额:58.0万元
-
批准年份:2019
-
负责人:高学文
-
依托单位:
变异链球菌small RNAs连接LuxS密度感应与生物膜形成的机制研究
-
批准号:81900988
-
项目类别:青年科学基金项目
-
资助金额:21.0万元
-
批准年份:2019
-
负责人:毛梦莹
-
依托单位:
肠道细菌关键small RNAs在克罗恩病发生发展中的功能和作用机制
-
批准号:31870821
-
项目类别:面上项目
-
资助金额:56.0万元
-
批准年份:2018
-
负责人:陈江宁
-
依托单位:
基于small RNA 测序技术解析鸽分泌鸽乳的分子机制
-
批准号:31802058
-
项目类别:青年科学基金项目
-
资助金额:26.0万元
-
批准年份:2018
-
负责人:麻慧
-
依托单位:
Small RNA介导的DNA甲基化调控的水稻草矮病毒致病机制
-
批准号:31772128
-
项目类别:面上项目
-
资助金额:60.0万元
-
批准年份:2017
-
负责人:吴建国
-
依托单位:
基于small RNA-seq的针灸治疗桥本甲状腺炎的免疫调控机制研究
-
批准号:81704176
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2017
-
负责人:赵继梦
-
依托单位:
水稻OsSGS3与OsHEN1调控small RNAs合成及其对抗病性的调节
-
批准号:91640114
-
项目类别:重大研究计划
-
资助金额:85.0万元
-
批准年份:2016
-
负责人:何祖华
-
依托单位: