课题基金 / 基金详情

Enabling Relational Reasoning in Multi-Agent Deep Reinforcement Learning

Enabling Relational Reasoning in Multi-Agent Deep Reinforcement Learning
在多智能体深度强化学习中实现关系推理
批准号:
2585630
负责人:
金额:
$0.0万
依托单位:
依托单位国家:
英国
项目类别:
Studentship
财政年份:
2021
资助国家:
英国
项目状态:
未结题
起止时间:
2021 至 --

项目摘要

项目成果

相似基金

相关文献

中文摘要
翻译
点击翻译按钮获取中文摘要
英文摘要
The aim of reinforcement learning is to teach an artificial agent how to take optimal sequential decisions in an uncertain environment to complete a task. Recent advancements in the field have leveraged deep learning methods as function approximators thus enabling complex applications in various areas, from gaming to bioinformatics. Despite these advances, most of the work has focused on the case of a single agent interacting with the environment. However, many real-world applications involve multiple cooperative agents taking joint decisions; some prominent examples include autonomous vehicles, manufacturing robotics and cyber-security bots. In such settings, inter-agent communication becomes essential to achieve collaborative behaviour, and recent developments in the fields have been concerned with facilitating the spontaneous emergence of communication protocols throughout the learning process. In this project, we will develop a modelling framework where, in addition to learning how to communicate, the agents can also develop the ability to perform relational reasoning, i.e. they'll be able to infer how the entities acting in the environment are related to one another and encode those relationships in order to improve the decision-making process. In doing so, we will draw heavily from the field of geometric deep learning where relational graph neural networks are currently employed to learn relational patterns from network-valued data. Our aim is to develop a unified relational reinforcement learning approach for multi agent systems that is both decentralised and scalable. Several applications of increasing complexity will be considered to showcase the potential use of our algorithms in real-world use cases.
期刊论文(0)
专著(0)
科研奖励(0)
会议论文
海外基金