CRII: Learning to simulate with small data
CRII: Learning to simulate with small data
批准号:
2153311
负责人:
Hua Wei
金额:
$17.5万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2022
资助国家:
美国
项目状态:
已结题
起止时间:
2022-07-01 至 2024-06-30
中文摘要
强化学习(RL)在围棋等一系列人工智能领域取得了巨大的成功。由于RL可以提供优化的策略来实现特定的目标,人们迫切希望使用先进的RL技术来解决现实世界的决策问题。尽管RL在人工智能领域取得了巨大的成功,但它在现实世界的应用中还没有表现出同样程度的成功,因为RL在很大程度上依赖于模拟,而且很少有一个好的模拟器来模拟真实世界的系统。与游戏等模拟环境不同,数据可能是无限的,而现实世界中的物理系统,如交通系统,情况正好相反:数据很小(即,稀疏且难以获取)。这就提出了一个重要的研究问题,即如何从小数据中构建逼真的模拟,以模拟复杂和随机的真实世界动态。该项目的解决方案可以帮助政策制定者在现实世界中实施之前选择更好的策略,极大地促进强化学习技术在现实世界中的采用,并有利于许多希望使用真实数据来更好地学习或理解真实世界物理系统的应用程序。该项目旨在通过研究数据挖掘算法来构建一个现实的交通模拟器,并提供利用小数据模拟真实世界模拟的解决方案。该项目将学习在不对真实世界的模型做出不切实际的假设的情况下进行模拟,并进一步学习小数据的真实世界设置。首先,该项目将尝试从真实世界的不完整和间接观察中学习数据驱动的模型。其次,该项目将寻求创新数据驱动的模型,以满足人类知识,因为人类知识可以指导我们学习一种不仅依赖于小数据和可能存在偏见的数据的模型。第三,该项目旨在利用现实世界中多方对数据驱动模型的影响。在数据挖掘过程中将开发新的机器学习技术。这一奖项反映了NSF的法定使命,并通过使用基金会的智力优势和更广泛的影响审查标准进行评估,被认为值得支持。
英文摘要
Reinforcement learning (RL) has shown great success in a series of artificial intelligence (AI) domains such as Go games. Since RL can provide optimized policies to achieve certain goals, people are eager to use advanced RL techniques to solve real-world decision-making problems. Despite its huge success in AI domains, RL has not yet shown the same degree of success for real-world applications because RL largely relies on simulation, and there is rarely a good simulator for real-world systems. Unlike simulated environments such as games where data could be unlimited, real-world physical systems like traffic systems are quite the opposite: data is small (i.e., sparse and hard to obtain). This raises important research questions about building realistic simulations from small data that can mimic complex and stochastic real-world dynamics. The solution from this project can help policymakers choose a better policy before implementing it in the real world, greatly facilitate the adoption of reinforcement learning techniques in the real world, and benefit many applications in which one would like to use real data to learn or understand real-world physical systems better.This project aims to build a realistic traffic simulator by investigating data mining algorithms and provides solutions toward mimicking real-world simulations with small data with applications to traffic simulations. This project will learn to simulate without making unrealistic assumptions on the real-world models and further learn with the real-world setting of small data. First, the project will try to learn data-driven models from incomplete and indirect observations from the real world. Second, this project will seek to innovate the data-driven model to meet with human knowledge, as human knowledge could guide us to learn a model that is not only relied on small and possibly biased data. Third, this project aims to leverage the influences from multiple parties in the real world for the data-driven model. New machine learning techniques will be developed in the data mining process.This award reflects NSF's statutory mission and has been deemed worthy of support through evaluation using the Foundation's intellectual merit and broader impacts review criteria.
期刊论文(3)
专著(0)
科研奖励(0)
会议论文
The Third Workshop on Data-driven Intelligent Transportation
第三届数据驱动智能交通研讨会
DOI:
10.1145/3511808.3557496
发表时间:
2022
期刊:
CIKM '22: Proceedings of the 31st ACM International Conference on Information & Knowledge Management
影响因子:
--
作者:
[Wei, Hua, Sheron, Guni, Wu, Cathy, Chawla, Sanjay, Li, Zhenhui]
通讯作者:
Li, Zhenhui
DOI:
10.3934/era.2023057
发表时间:
2022
期刊:
Electronic Research Archive
影响因子:
0.8
作者:
[Longchao Da;Hua Wei]
通讯作者:
Longchao Da;Hua Wei
DOI:
10.1145/3534678.3539236
发表时间:
2022-08
期刊:
Proceedings of the 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining
影响因子:
--
作者:
[Xiaoliang Lei;Hao Mei;Bin Shi;Hua Wei]
通讯作者:
Xiaoliang Lei;Hao Mei;Bin Shi;Hua Wei
国内基金
海外基金
登录
查看更多内容
Scalable Learning and Optimization: High-dimensional Models and Online Decision-Making Strategies for Big Data Analysis
-
批准号:--
-
项目类别:合作创新研究团队
-
资助金额:--
-
批准年份:2024
-
负责人:姚韬
-
依托单位:
Understanding structural evolution of galaxies with machine learning
-
批准号:
-
项目类别:省市级项目
-
资助金额:10.0万元
-
批准年份:2022
-
负责人:Nicola Rosario Napolitano
-
依托单位:
煤矿安全人机混合群智感知任务的约束动态多目标Q-learning进化分配
-
批准号:--
-
项目类别:青年科学基金项目
-
资助金额:30万元
-
批准年份:2022
-
负责人:吉建娇
-
依托单位:
基于领弹失效考量的智能弹药编队短时在线Q-learning协同控制机理
-
批准号:62003314
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2020
-
负责人:沈剑
-
依托单位:
集成上下文张量分解的e-learning资源推荐方法研究
-
批准号:61902016
-
项目类别:青年科学基金项目
-
资助金额:24.0万元
-
批准年份:2019
-
负责人:万珊珊
-
依托单位:
具有时序迁移能力的Spiking-Transfer learning (脉冲-迁移学习)方法研究
-
批准号:61806040
-
项目类别:青年科学基金项目
-
资助金额:20.0万元
-
批准年份:2018
-
负责人:解修蕊
-
依托单位:
基于Deep-learning的三江源区冰川监测动态识别技术研究
-
批准号:51769027
-
项目类别:地区科学基金项目
-
资助金额:38.0万元
-
批准年份:2017
-
负责人:张大奇
-
依托单位:
具有时序处理能力的Spiking-Deep Learning(脉冲深度学习)方法研究
-
批准号:61573081
-
项目类别:面上项目
-
资助金额:64.0万元
-
批准年份:2015
-
负责人:屈鸿
-
依托单位:
基于有向超图的大型个性化e-learning学习过程模型的自动生成与优化
-
批准号:61572533
-
项目类别:面上项目
-
资助金额:66.0万元
-
批准年份:2015
-
负责人:孙雪冬
-
依托单位:
E-Learning中学习者情感补偿方法的研究
-
批准号:61402392
-
项目类别:青年科学基金项目
-
资助金额:26.0万元
-
批准年份:2014
-
负责人:秦继伟
-
依托单位: