Using Wizard-of-Oz simulations to bootstrap Reinforcement - Learning based dialog management systems
Using Wizard-of-Oz simulations to bootstrap Reinforcement - Learning based dialog management systems
复制标题
使用绿野仙踪模拟来引导强化 - 基于学习的对话管理系统
DOI:
--
复制
发表时间:
2003
期刊:
影响因子:
--
通讯作者:
S. Young
中科院分区:
文献类型:
--
作者:
J. Williams;S. Young
This paper describes a method for “bootstrapping” a Reinforcement Learningbased dialog manager using a Wizard-ofOz trial. The state space and action set are discovered through the annotation, and an initial policy is generated using a Supervised Learning algorithm. The method is tested and shown to create an initial policy which performs significantly better and with less effort than a handcrafted policy, and can be generated using a small number of dialogs.