课题基金 / 基金详情

NSF-AoF: RI: Small: Safe Reinforcement Learning in Non-Stationary Environments With Fast Adaptation and Disturbance Prediction

NSF-AoF: RI: Small: Safe Reinforcement Learning in Non-Stationary Environments With Fast Adaptation and Disturbance Prediction
NSF-AoF:RI:小型:具有快速适应和干扰预测功能的非平稳环境中的安全强化学习
批准号:
2133656
负责人:
Naira Hovakimyan
金额:
$50.0万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2021
资助国家:
美国
项目状态:
已结题
起止时间:
2021-09-01 至 2024-08-31
关键词:

项目摘要

项目成果

Naira Hovakimyan的其他基金

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
Reinforcement learning (RL) has shown impressive performance in the control of complex robotic systems for various tasks such as locomotion, manipulation, and playing sports, e.g., table tennis. Reinforcement learning enables a robot to autonomously discover an optimal behavior through trial-and-error interactions with its environment. However, the environmental perturbations could easily cause a behavior policy trained in an old environment to fail in a perturbed environment. The failure is unacceptable for safety-critical robotic systems such as self-driving cars, drones, flying taxies and construction machines. Existing robust methods try to consider all scenarios during the training phase and seek a fixed policy, leading to conservative behaviors. Existing adaptive methods try to update their behavior policies in the perturbed environment, but will only do that after the robot has “felt a difference” through its interaction with the environment. In contrast, a human could leverage his/her perception for prediction in the new environment and adjust his/her behavior accordingly even before interacting with it. In light of these conditions, this project envisions a new framework for safe and efficient RL in the presence of environmental changes leveraging fast adaptation and perception-based prediction. The framework will enable robotic and autonomous systems robustly and safely operate, learn and adapt in the real world. This project relies on the following thrusts: i) hybrid RL for safe and efficient policy updates, ii) robust adaptive control with safety guarantees; iii) vision-based disturbance prediction. More specifically, the project will develop robust adaptive control algorithms that ensure that the executed trajectory of a robot remains safe in the presence of disturbances induced by environmental changes. It will spur hybrid model-free/model-based RL algorithms that are capable of efficiently and safely updating the behavior policies with the help of the control algorithms. The project will advance novel methodologies for predicting the key parameters of the disturbances (e.g., the weight of a package) directly from the image observations, leading to new scalable methods for efficiently learning the mathematical model of the disturbances with quantified error bounds. All the ingredients will be holistically integrated to build a framework to enable robots to safely, robustly, and efficiently operate and adapt in real-world environments. Aerial and ground vehicles will be used for experimental validation.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(9)
专著(0)
科研奖励(0)
会议论文
DOI: 10.1109/lcsys.2023.3293765
发表时间: 2023
期刊: IEEE Control Systems Letters
影响因子: 3
作者: [Pan Zhao;R. Ghabcheloo;Yikun Cheng;Hossein Abdi;N. Hovakimyan]
通讯作者: Pan Zhao;R. Ghabcheloo;Yikun Cheng;Hossein Abdi;N. Hovakimyan
DOI: 10.1109/lra.2022.3153712
发表时间: 2021-09
期刊: IEEE Robotics and Automation Letters
影响因子: 5.2
作者: [Pan Zhao;Arun Lakshmanan;K. Ackerman;Aditya Gahlawat;M. Pavone;N. Hovakimyan]
通讯作者: Pan Zhao;Arun Lakshmanan;K. Ackerman;Aditya Gahlawat;M. Pavone;N. Hovakimyan
DOI: 10.48550/arxiv.2212.03194
发表时间: 2022-12
期刊:
影响因子: --
作者: [Sheng Cheng;Lin Song;Minkyung Kim;Shenlong Wang;N. Hovakimyan]
通讯作者: Sheng Cheng;Lin Song;Minkyung Kim;Shenlong Wang;N. Hovakimyan
DOI: --
发表时间: 2022-11
期刊:
影响因子: --
作者: [Yikun Cheng;Pan Zhao;N. Hovakimyan]
通讯作者: Yikun Cheng;Pan Zhao;N. Hovakimyan
6
    Collaborative Research: SLES: Guaranteed Tubes for Safe Learning across Autonomy Architectures
    Distributionally Robust Adaptive Control: Enabling Safe and Robust Reinforcement Learning
    NRI: INT: COLLAB: Synergetic Drone Delivery Network in Metropolis
    CPS: Medium: Collaborative Research: Against Coordinated Cyber and Physical Attacks: Unified Theory and Technologies
    海外基金