Service Function Chain Placement in Cloud Data Center Networks: A Cooperative Multi-agent Reinforcement Learning Approach
Service Function Chain Placement in Cloud Data Center Networks: A Cooperative Multi-agent Reinforcement Learning Approach
复制标题
DOI:
10.1007/978-3-031-23141-4_22
复制
发表时间:
2022
期刊:
影响因子:
--
通讯作者:
Lynn Gao;Yutian Chen;Bin Tang
中科院分区:
文献类型:
--
作者:
Lynn Gao;Yutian Chen;Bin Tang
Service function chaining (SFC), consisting of a sequence of virtual network functions (VNFs) (i.e., firewalls and load balancers), is an effective service provision technique in modern data center networks. By requiring cloud user traffic to traverse the VNFs in order, SFC improves the security and performance of the cloud user applications. In this paper, we study how to place an SFC inside a data center to minimize the network traffic of the virtual machine (VM) communication. We take a cooperative multi-agent reinforcement learning approach, wherein multiple agents collaboratively figure out the traffic-efficient route for the VM communication.Underlying the SFC placement is a fundamental graph-theoretical problem called thek-stroll problem. Given a weighted graphG(V,E), two nodess, \documentclass[12pt]{minimal} \usepackage{amsmath} \usepackage{wasysym} \usepackage{amsfonts} \usepackage{amssymb} \usepackage{amsbsy} \usepackage{mathrsfs} \usepackage{upgreek} \setlength{\oddsidemargin}{-69pt} \begin{document}$$t \in V$$\end{document}, and an integerk, thek-stroll problem is to find the shortest path fromstotthat visits at leastkother nodes in the graph. Our work is the first to take a multi-agent learning approach to solvek-stroll problem. We compare our learning algorithm with an optimal and exhaustive algorithm and an existing dynamic programming(DP)-based heuristic algorithm. We show that our learning algorithm, although lacking the complete knowledge of the network assumed by existing research, delivers comparable or even better VM communication time while taking two orders of magnitude of less execution time.