Understanding and Combining Sequential Decision Making Methods
Understanding and Combining Sequential Decision Making Methods
批准号:
RGPIN-2021-03099
负责人:
Valenzano, Richard
金额:
$1.75万
依托单位:
依托单位国家:
加拿大
项目类别:
Discovery Grants Program - Individual
财政年份:
2022
资助国家:
加拿大
项目状态:
已结题
起止时间:
2022-01-01 至 2023-12-31
中文摘要
点击翻译按钮获取中文摘要
英文摘要
If a robot is to successfully navigate from one location to another, it must identify an appropriate sequence of movements to get there. Similarly, a chess-playing program must select a sequence of "good" moves in response to its opponent if the program is to win. These are examples of sequential decision making tasks, which require that an autonomous agent make a sequence of effective decisions about how to act in order to achieve some desired objective. Sequential decision making is a fundamental task in the field of Artificial Intelligence, and it is becoming even more important as autonomous systems become a larger part of daily life. It has been examined by several different research communities, each considering different settings and using different solution methods. This includes reinforcement learning, automated planning, and game-playing systems. The long-term goal of my research agenda is to develop the underlying principles and techniques needed for effective sequential decision making. I will pursue two main lines of inquiry. The first is to improve our understanding of different sequential decision making methods --- specifically reinforcement learning, automated planning, and game-playing approaches --- with the objective of making it easier to develop systems to solve new or more complex sequential decision making tasks. This includes an examination of how different design decisions affect system performance, and an investigation into what properties of a problem lead to one technique being more effective than another. The second main line of research is in understanding how different sequential decision making methods can be used to improve and complement each other. AlphaGo is an excellent example of this idea, as this system used reinforcement learning to generate an evaluation function, and then used this function to guide a planning method during play. This general paradigm is a very powerful one, and better understanding when it is most appropriate or how to best combine other such methods should advance our ability to make systems with strong sequential decision making capabilities. Sequential decision making methods have been used in a variety of applications including DNA sequence alignment, underground sewer placement for new subdivisions, model-based diagnosis of faulty systems, and robotics. However, it is still quite complex to design a system that uses sequential decision making approaches for some new given task, due to the significant number of design decisions that need to be made. Having a better understanding of when different methods are most effective will help speed up this process, while identifying new ways of combining the various methods for sequential decision making will increase the size and complexity of problems that can be solved.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
Understanding and Combining Sequential Decision Making Methods
-
批准号:DGECR-2021-00222
-
项目类别:Discovery Launch Supplement
-
资助金额:$0.91万
-
财政年份:2021
-
负责人:Valenzano, Richard
-
依托单位:
Understanding and Combining Sequential Decision Making Methods
-
批准号:RGPIN-2021-03099
-
项目类别:Discovery Grants Program - Individual
-
资助金额:$1.75万
-
财政年份:2021
-
负责人:Valenzano, Richard
-
依托单位:
Dynamic weighting in single-agent search algorithms
-
批准号:346815-2009
-
项目类别:Postgraduate Scholarships - Doctoral
-
资助金额:$1.53万
-
财政年份:2010
-
负责人:Valenzano, Richard
-
依托单位:
Dynamic weighting in single-agent search algorithms
-
批准号:346815-2009
-
项目类别:Postgraduate Scholarships - Doctoral
-
资助金额:$1.53万
-
财政年份:2009
-
负责人:Valenzano, Richard
-
依托单位:
Postgraduate scholarship application for Richard Valenzano with regards to an undecided research project
-
批准号:346815-2008
-
项目类别:Postgraduate Scholarships - Master's
-
资助金额:$1.26万
-
财政年份:2008
-
负责人:Valenzano, Richard
-
依托单位:
Postgraduate scholarship application for Richard Valenzano with regards to an undecided research project
-
批准号:346815-2007
-
项目类别:Alexander Graham Bell Canada Graduate Scholarships - Master's
-
资助金额:$1.27万
-
财政年份:2007
-
负责人:Valenzano, Richard
-
依托单位:
NAViGaTor - Extending interactive visualization of protein interaction networks
-
批准号:353144-2007
-
项目类别:University Undergraduate Student Research Awards
-
资助金额:$0.33万
-
财政年份:2007
-
负责人:Valenzano, Richard
-
依托单位:
海外基金