Collaborative Research: SLES: Guaranteed Tubes for Safe Learning across Autonomy Architectures
Collaborative Research: SLES: Guaranteed Tubes for Safe Learning across Autonomy Architectures
批准号:
2331878
负责人:
Naira Hovakimyan
金额:
$96.91万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2024
资助国家:
美国
项目状态:
未结题
起止时间:
2024-01-01 至 2027-12-31
中文摘要
自主运行的自治系统(例如,自动汽车和送货无人机)暴露出取决于任务和环境的复杂性的安全挑战。虽然这些系统已经证明了自己学习和适应的能力,但确保这些系统的安全性并不是小事。安全可能由于各种因素而受到威胁,包括但不限于意外变化、恶劣天气条件或未知障碍物。这些系统的机载学习解决方案主要用于非关键情况,如计算机游戏,其中安全问题最小。该提案旨在解决各种应用场景中学习支持系统的端到端安全性的迫切需求,例如,城市空中交通中的自动驾驶汽车和飞行器。我们提出了一种新的解决方案,称为“数据启用单纯形”或“DeSimplex”。DeSimplex建立在坚实的数学原理和系统方法之上,用于收集数据并将数据用于系统的性能改进。它提供了一个可以被证明是安全的框架,并允许学习型系统即使在面临极端事件或环境危害时也能进行调整并表现良好。拟议的工作对于更广泛的应用至关重要,这些应用涉及自主系统在不可预测、要求严格的物理环境中安全、高效地运行,包括自主汽车和3D城市空中移动的飞行器。拟议的工作为推进端到端学习支持系统中的安全理解奠定了基础,这是网络物理系统、机器人技术和机器学习的基础问题。 我们的目标是追求以下两个相互关联的研究方向:(i)通过可靠的不确定性量化方法提高高性能自治,以确保数据驱动的适应性和不确定性的精确测量;(ii)开发高保证的自治架构,并建立可验证的可观测性和可控性的切换规则。将开发一个框架,结合高性能和高保证自主性的优势,促进适应性学习,准确的不确定性量化和可验证的安全措施。将开发新的方法,用于(i)基于策略的闭环学习以提高性能,(ii)可靠的不确定性量化以提供数据驱动的适应性,以及(iii)系统级的可验证可观测性和可控性。将采用系统的双策略方法来安全收集数据,以实现拟议的高性能自主性,从而协调数据的三个理想属性:安全性、策略性和闭环。建议的框架将在一个严格的程序中验证,从模块化模拟测试到集成和部署在真实的空中和地面车辆上。这项研究得到了美国国家科学基金会和开放慈善机构之间的合作伙伴关系的支持。这个奖项反映了NSF的法定使命,并被认为是值得通过使用基金会的智力价值和更广泛的影响审查标准进行评估的支持。
英文摘要
The autonomous systems that operate by themselves (e.g., autonomous cars and delivery drones) expose safety challenges dependent upon the complexity of the missions and environments. Although such systems have demonstrated the ability to learn and adapt on their own, ensuring the safety of these systems is not trivial. Safety can be endangered due to various factors, including but not limited to unexpected changes, inclement weather conditions, or unknown obstacles. The onboard learning solutions of these systems have mostly been used in non-critical situations like computer games, where safety concerns are minimal. This proposal aims to address the urgent need for end-to-end safety in learning-enabled systems across various application scenarios, e.g., self-driving cars and flying vehicles in urban air mobility. We propose a novel solution called "Data-enabled Simplex" or "DeSimplex.” DeSimplex is built on solid mathematical principles and systematic methods for collecting data and using the data for the system’s performance improvement. It provides a framework that can be proven to be safe and allows learning-enabled systems to adjust and perform well even when faced with extreme events, or environmental hazards. The proposed work is crucial for wider applications that involve the safe and efficient operation of autonomous systems in unpredictable, demanding physical environments, including autonomous cars and flying vehicles of 3D urban air mobility.The proposed work lays the groundwork for advancing the comprehension of safety in end-to-end learning-enabled systems, a foundational problem in cyber-physical systems, robotics, and machine learning. We aim to pursue the following two interconnected research thrusts: (i) improving high-performance autonomy with reliable uncertainty quantification methods to ensure data-driven adaptability and precise measurement of uncertainties and (ii) developing high-assurance autonomy architectures and establishing switching rules for verifiable observability and controllability. A framework will be developed that combines the strengths of high-performance and high-assurance autonomy, facilitating adaptive learning, accurate uncertainty quantification, and verifiable safety measures. Novel methods will be developed for (i) on-policy, closed-loop learning to boost performance, (ii) reliable uncertainty quantification to provide data-driven adaptability, and (iii) verifiable observability and controllability at the system level. A systematic, dual-strategy approach will be pursued for safe data collection for the proposed high-performance autonomy to reconcile the three desired properties for data: safety, on-policy, and closed-loop. The proposed framework will be validated in a rigorous procedure from modular simulation testing to integration and deployment on real aerial and ground vehicles.This research is supported by a partnership between the National Science Foundation and Open Philanthropy.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Distributionally Robust Adaptive Control: Enabling Safe and Robust Reinforcement Learning
-
批准号:2135925
-
项目类别:Standard Grant
-
资助金额:$37.5万
-
财政年份:2022
-
负责人:Naira Hovakimyan
-
依托单位:
NSF-AoF: RI: Small: Safe Reinforcement Learning in Non-Stationary Environments With Fast Adaptation and Disturbance Prediction
-
批准号:2133656
-
项目类别:Standard Grant
-
资助金额:$50.0万
-
财政年份:2021
-
负责人:Naira Hovakimyan
-
依托单位:
NRI: INT: COLLAB: Synergetic Drone Delivery Network in Metropolis
-
批准号:1830639
-
项目类别:Standard Grant
-
资助金额:$103.77万
-
财政年份:2018
-
负责人:Naira Hovakimyan
-
依托单位:
CPS: Medium: Collaborative Research: Against Coordinated Cyber and Physical Attacks: Unified Theory and Technologies
-
批准号:1739732
-
项目类别:Standard Grant
-
资助金额:$70.0万
-
财政年份:2017
-
负责人:Naira Hovakimyan
-
依托单位:
NRI: Collaborative Research: ASPIRE: Automation Supporting Prolonged Independent Residence for the Elderly
-
批准号:1528036
-
项目类别:Standard Grant
-
资助金额:$129.59万
-
财政年份:2015
-
负责人:Naira Hovakimyan
-
依托单位:
EAGER: Human centered robotic system design
-
批准号:1548409
-
项目类别:Standard Grant
-
资助金额:$30.0万
-
财政年份:2015
-
负责人:Naira Hovakimyan
-
依托单位:
国内基金
海外基金
登录
查看更多内容
Research on Quantum Field Theory without a Lagrangian Description
-
批准号:24ZR1403900
-
项目类别:省市级项目
-
资助金额:--
-
批准年份:2024
-
负责人:SATOSHI NAWATA
-
依托单位:
Cell Research
-
批准号:31224802
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2012
-
负责人:程磊
-
依托单位:
Cell Research
-
批准号:31024804
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2010
-
负责人:程磊
-
依托单位:
Cell Research (细胞研究)
-
批准号:30824808
-
项目类别:专项基金项目
-
资助金额:24.0万元
-
批准年份:2008
-
负责人:张爱兰
-
依托单位:
Research on the Rapid Growth Mechanism of KDP Crystal
-
批准号:10774081
-
项目类别:面上项目
-
资助金额:45.0万元
-
批准年份:2007
-
负责人:滕冰
-
依托单位: