EAGER: Real-Time: Formal Reinforcement Learning Methods for the Design of Safety-critical Autonomous Systems
EAGER: Real-Time: Formal Reinforcement Learning Methods for the Design of Safety-critical Autonomous Systems
批准号:
1839842
负责人:
Rahul Jain
金额:
$28.61万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2019
资助国家:
美国
项目状态:
已结题
起止时间:
2019-04-01 至 2022-03-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
This EArly-Concept Grant for Exploratory Research (EAGER) project takes a clean-slate first-principles approach to the design of safety-critical autonomous systems by integrating formal methods and reinforcement learning from data. Several recent high-profile traffic incidents involving semi-autonomous vehicles have raised questions about whether current artificial intelligence (AI)-centered methods can ever lead us to Level 4 or 5 autonomy, i.e., to the realization of fully-autonomous vehicles with performance equivalent to a human driver in all driving scenarios. On the other hand, approaches rooted in formal methods for verification and synthesis can provide safety guarantees but have difficulty in efficiently reasoning about uncertainty and the correctness of data-driven models. This project will combine these two, seemingly incompatible, paradigms for designing autonomous systems. It will use model-free reinforcement learning algorithms to learn from semi-autonomous vehicle driving data. It will adopt model-based methods for system design, verification, and synthesis to offer provably safe operation in highly uncertain scenarios. An AutoDrive testbed will be set up where human driving data from scaled vehicular models will be leveraged to infer safe control policies using imitation and inverse reinforcement learning algorithms. The research is relevant to the science of intelligent autonomous transportation systems with significant societal implications. The experimental testbed will be used to provide hands-on research experience to undergraduate students and for K-12 outreach efforts.In particular, the project will develop a framework for optimal control synthesis for safety and performance specification expressed in signal temporal logic. It will then incorporate vehicular and pedestrian kinematics in non-deterministic/probabilistic transition models specified via probabilistic computation tree logic. Finally, it will develop formal reinforcement learning methods for partially observed dynamic models subject to safety specifications and complex temporal goals by learning from traces of safe human drivers. One key technical contribution of the project will be development of new formal reinforcement learning methods that may be useful in a broad array of applications wherein we must synthesize optimal controllers that satisfy certain safety specifications by learning from data.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(2)
专著(0)
科研奖励(0)
会议论文
Model-Free Reinforcement Learning for Optimal Control of Markov Decision Processes Under Signal Temporal Logic Specifications
信号时序逻辑规范下马尔可夫决策过程最优控制的无模型强化学习
DOI:
10.1109/cdc45484.2021.9683444
发表时间:
2021
期刊:
2021 60th IEEE Conference on Decision and Control (CDC
影响因子:
--
作者:
[Kalagarla, Krishna C., Jain, Rahul, Nuzzo, Pierluigi]
通讯作者:
Nuzzo, Pierluigi
Model-free Reinforcement Learning in Infinite-horizon Average-reward Markov Decision Processes”, Proc. ICML 2020. (arXiv:1910.07072)
无限视野平均奖励马尔可夫决策过程中的无模型强化学习,Proc。
DOI:
--
发表时间:
2020
期刊:
Proceedings of the 37th International Conference on Machine Learning
影响因子:
--
作者:
[Chen-Yu Wei, Mehdi Jafarnia-Jahromi]
通讯作者:
Chen-Yu Wei, Mehdi Jafarnia-Jahromi
Online Learning-based Real-time Control of Unknown Autonomous Systems
-
批准号:1810447
-
项目类别:Standard Grant
-
资助金额:$33.0万
-
财政年份:2018
-
负责人:Rahul Jain
-
依托单位:
AF: Small: A New Approach to Analysis and Design of Algorithms for Stochastic Control and Optimization
-
批准号:1817212
-
项目类别:Standard Grant
-
资助金额:$40.0万
-
财政年份:2018
-
负责人:Rahul Jain
-
依托单位:
Collaborative Research: Smarter Markets for a Smarter Grid: Pricing Randomness, Flexibility and Risk
-
批准号:1611574
-
项目类别:Standard Grant
-
资助金额:$22.5万
-
财政年份:2016
-
负责人:Rahul Jain
-
依托单位:
CAREER: Network Economics: Theory and Architectures for Incentive-engineered Networks
-
批准号:0954116
-
项目类别:Continuing Grant
-
资助金额:$42.5万
-
财政年份:2010
-
负责人:Rahul Jain
-
依托单位:
NetSE: Small: Cooperation and Incentives in Communication and Social Networks
-
批准号:0917410
-
项目类别:Continuing Grant
-
资助金额:$44.25万
-
财政年份:2009
-
负责人:Rahul Jain
-
依托单位:
国内基金
海外基金
Immuno-Real Time PCR法精确定量血清MG7抗原及在早期胃癌预警中的价值
-
批准号:30600737
-
项目类别:青年科学基金项目
-
资助金额:22.0万元
-
批准年份:2006
-
负责人:陈峥
-
依托单位:
无色ReAl3(BO3)4(Re=Y,Lu)系列晶体紫外倍频性能与器件研究
-
批准号:60608018
-
项目类别:青年科学基金项目
-
资助金额:28.0万元
-
批准年份:2006
-
负责人:叶宁
-
依托单位: