RI: Medium: Building Flexible, Robust, and Autonomous Agents
RI: Medium: Building Flexible, Robust, and Autonomous Agents
批准号:
0905146
负责人:
Satinder Baveja
金额:
$120.0万
依托单位国家:
美国
项目类别:
Standard Grant
财政年份:
2009
资助国家:
美国
项目状态:
已结题
起止时间:
2009-07-01 至 2013-09-30
中文摘要
点击翻译按钮获取中文摘要
英文摘要
This project is developing computational agents that operate for extended periods of time in rich and dynamic environments, and achieve mastery of many aspects of their environments without task-specific programming. To accomplish these goals, research is exploring a space of cognitive architectures that incorporate four fundamental features of real neural circuitry: (1) reinforcing behaviors that lead to intrinsic rewards, (2) executing and learning over mental, as well as, motor actions, (3) extracting regularities in mental representations, whether derived from perception or cognitive operations, and (4) continuously encoding and retrieving episodic memories of past events. A software framework called Storm facilitates this exploration by enabling the integration of independent functional subsystems, allowing researchers to easily plug in and remove different subsystems in order to assess their impact on the overall behavior of the system. Cognitive architectures are being tested by exposing them to a wide variety of novel environments with unpredictable (and non-repeatable) extrinsic rewards, but in which many actions could lead to intrinsic rewards (e.g., surprise). To assess flexibility, an automated environment generator exposes agents to environments that are unknown in advance to the artificial agent or human researcher. To assess robustness, cognitive systems are being exposed to many variants of the same environment to ensure that the systems can learn from past experience and generalize when appropriate. And to assess autonomy, systems' must operate effectively for extended periods of time in a dynamic environment. In the longer term, flexible and robust cognitive architectures being devloped under this research will have application as the 'brains' of robotic and software systems in emergency, miltary, and a wide variety of other societal and service realms. This award is funded under the American Recovery and Reinvestment Act of 2009 (Public Law 111-5).
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
RI: Small: Combining Reinforcement Learning and Deep Learning Methods to Address High-Dimensional Perception, Partial Observability and Delayed Reward
-
批准号:1526059
-
项目类别:Standard Grant
-
资助金额:$49.99万
-
财政年份:2015
-
负责人:Satinder Baveja
-
依托单位:
RI: Small: Reinforcement Learning with Predictive State Representations
-
批准号:1319365
-
项目类别:Continuing Grant
-
资助金额:$45.0万
-
财政年份:2013
-
负责人:Satinder Baveja
-
依托单位:
EAGER: On the Optimal Rewards Problem
-
批准号:1148668
-
项目类别:Standard Grant
-
资助金额:$20.0万
-
财政年份:2011
-
负责人:Satinder Baveja
-
依托单位:
SHB: Medium: Collaborative Research: Novel Computational Techniques for Cardiovascular Risk Stratification
-
批准号:1064948
-
项目类别:Standard Grant
-
资助金额:$56.24万
-
财政年份:2011
-
负责人:Satinder Baveja
-
依托单位:
Flexible State Representations in Reinforcement Learning
-
批准号:0413004
-
项目类别:Continuing Grant
-
资助金额:$0.0万
-
财政年份:2005
-
负责人:Satinder Baveja
-
依托单位:
Collaborative Research: Intrinsically Motivated Learning in Artificial Agents
-
批准号:0432027
-
项目类别:Continuing Grant
-
资助金额:$15.0万
-
财政年份:2004
-
负责人:Satinder Baveja
-
依托单位:
Exploiting Structure in Reinforcement Learning Problems
-
批准号:9711753
-
项目类别:Continuing Grant
-
资助金额:$22.97万
-
财政年份:1997
-
负责人:Satinder Baveja
-
依托单位:
海外基金