课题基金 / 基金详情

Open Domain Statistical Spoken Dialogue Systems

Open Domain Statistical Spoken Dialogue Systems
开放域统计口语对话系统
批准号:
EP/M018946/1
负责人:
Stephen Young
金额:
$76.89万
依托单位:
依托单位国家:
英国
项目类别:
Research Grant
财政年份:
2015
资助国家:
英国
项目状态:
已结题
起止时间:
2015 至 --

项目摘要

项目成果

Stephen Young的其他基金

相似基金

相关文献

中文摘要
翻译
口语对话系统(SDS)包括建立有效的人机界面所需的技术,主要依赖于语音。到目前为止,它们主要部署在基于电话的呼叫中心应用程序中,如银行,账单查询和旅行信息,它们是使用手工制作的规则构建的。最近推出的Apple Siri和Google Now已经将基于语音的界面推向主流。这些虚拟个人助理(VPA)提供了彻底改变我们与机器交互方式的潜力,它们为正确控制和管理新兴的物联网开辟了道路--物联网是快速增长的智能设备网络,缺乏任何形式的传统用户界面。然而,当前的个人助理是使用与有限域口语对话系统相同的技术构建的。他们是不能够维持会话对话,除了在选定的有限的域,他们已经明确编程处理,最近的工作统计SDS已经表明,它不仅是可能的,这样一个系统,以适应和提高性能的域内,它已经被设计,但它也有可能为系统自动扩展其覆盖范围,包括新的,迄今未见的概念。这表明,应该有可能建立在有限领域统计SDS的发展取得的进展,设计一个全新的形式的口语对话系统(因此VPA),能够扩展和适应使用,以涵盖更广泛的会话主题。这样一个系统的设计是本次研究提案的重点,其关键思想是将最新的统计对话技术整合到一个覆盖面广的知识图谱中(例如freebase),它不仅包含关于实体的本体论信息,还包含可以应用于这些实体的操作(例如,查找航班信息、预订酒店房间、购买电子书等)。能够解释和响应每个可想到的用户请求的单个单片口语对话系统的实现是根本不可行的。因此,而不是简单地试图扩大现有SDS的覆盖范围,提出了一种新的分布式系统架构,具有三个关键特征:1。SDS的三个基本组件(语义解码器、对话管理器和响应生成器)分布在知识图上。本质上,图中的每个节点都有能力识别何时被引用,并有能力做出适当的响应。当用户说话时,所有的语义解码器都在听,基于解码器输出的激活级别,主题跟踪器识别哪个概念是焦点,并激活其对话策略。所有组件都是统计的,使得它们能够使用无监督的自适应来自动地在线自适应。通过确保类层次结构中的顶层节点具有经过良好训练的组件来管理数据稀疏性。最初,较低级别的更专业的概念只是从它们的超类继承所需的统计模型。当系统与用户交互并收集更多数据时,较低级别的组件会获得足够的数据来训练自己的专用统计模型。最终结果是系统不断在线学习。它以有限和呆板的对话风格开始,但使用得越多,它就变得越流畅,随着用户探索新主题,系统学会适应和扩展其处理这些新主题的能力。由于许多用户可以同时使用该系统,因此学习速度可以很快,并且能够适应底层数据的实时更新,所有这些都是虚拟个人助理必须具有的真正有用的特征。
英文摘要
Spoken Dialogue Systems (SDS) encompass the technologies required to build effective man-machine interfaces which depend primarily on voice. To date they have mostly been deployed in telephone-based call centre applications such as banking, billing queries and travel information and they are built using hand-crafted rules.The recent introduction of Apple Siri and Google Now has moved voice-based interfaces into the main-stream. These virtual personal assistants (VPAs) offer the potential to revolutionise the way we interact with machines, and they open the way to properly control and manage the emerging Internet of Things - the rapidly growing network of smart devices which lack any form of conventional user interface. However, current personal assistants are built using the same technology as limited domain spoken dialogue systems. They are not capable of sustaining conversational dialogues except within the selected limited domains which they have been explicitly programmed to handle.Very recent work on statistical SDS has demonstrated that it is not only possible for such a system to adapt and improve performance within the domain for which it has been designed but it is also possible for the system to automatically extend its coverage to include new, hitherto unseen concepts. This suggests that it should be possible to build on the progress achieved in the development of limited domain statistical SDS to design a radically new form of spoken dialogue system (and hence VPA) which is able to extend and adapt with use to cover an ever-wider range of conversational topics. The design of such a system is the focus of this research proposal.The key idea is to integrate the latest statistical dialogue technology into a wide coverage knowledge graph (such as freebase) which contains not only ontological information about entities but also the operations that can be applied to those entities (e.g. find flight information, book a hotel room, buy an ebook, etc. ).The implementation of a single monolithic spoken dialogue system capable of interpreting and responding to every conceivable user request is simply not practicable. Hence, rather than simply trying to broaden the coverage of existing SDS, a novel distributed system architecture is proposed with three key features:1. the three essential components of an SDS (semantic decoder, dialogue manager and response generator) are distributed across the knowledge-graph. In essence, every node in the graph has the capability to recognise when it is being referred to and have the capability to respond appropriately.2. when the user speaks, all semantic decoders are listening, based on the activation levels of the decoder outputs, a topic tracker identifies which concept is in focus and activates its dialogue policy.3. all components are statistical enabling them to be adapted automatically on-line using unsupervised adaptation. Data sparsity is managed by ensuring that the top level nodes in the class hierarchy have well-trained components. Initially, lower level more specialised concepts simply inherit the required statistical models from their super-classes. As the system interacts with users and more data is collected, lower level components acquire sufficient data to train their own dedicated statistical models.The end result is a system that continually learns on-line. It starts with a limited and stilted conversational style, but the more it is used, the more fluent it becomes, and as users explore new topics, the system learns to adapt and extend its capability to handle those new topics. Since many users can be using the system simultaneously, learning can be fast and capable of accommodating live updates of the underlying data, all of which are characteristics that a virtual personal assistant must have to be genuinely useful.
期刊论文(10)
专著(0)
科研奖励(0)
会议论文
DOI: 10.18653/v1/n18-2112
发表时间: 2018-03
期刊: ArXiv
影响因子: --
作者: [I. Casanueva;Paweł Budzianowski;Pei-hao Su;Stefan Ultes;L. Rojas-Barahona;Bo-Hsiang Tseng;M. Gašić]
通讯作者: I. Casanueva;Paweł Budzianowski;Pei-hao Su;Stefan Ultes;L. Rojas-Barahona;Bo-Hsiang Tseng;M. Gašić
DOI: 10.18653/v1/w18-5038
发表时间: 2018-07
期刊:
影响因子: --
作者: [I. Casanueva;Paweł Budzianowski;Stefan Ultes;Florian Kreyssig;Bo-Hsiang Tseng;Yen-Chen Wu;Milica Gasic]
通讯作者: I. Casanueva;Paweł Budzianowski;Stefan Ultes;Florian Kreyssig;Bo-Hsiang Tseng;Yen-Chen Wu;Milica Gasic
Distributed dialogue policies for multi-domain statistical dialogue management
多领域统计对话管理的分布式对话策略
DOI: 10.1109/icassp.2015.7178997
发表时间: 2015
期刊:
影响因子: --
作者: [Gasic M]
通讯作者: Gasic M
Exploiting Sentence and Context Representations in Deep Neural Models for Spoken Language Understanding
利用深度神经模型中的句子和上下文表示进行口语理解
DOI: 10.48550/arxiv.1610.04120
发表时间: 2016
期刊:
影响因子: --
作者: [Barahona L]
通讯作者: Barahona L
共 7 条
    Doctoral Dissertation Research: The Economic and Environmental Tradeoffs of Concrete Construction in Urban Settings
    • 批准号:
      2113938
    • 项目类别:
      Standard Grant
    • 资助金额:
      $2.02万
    • 财政年份:
      2021
    • 负责人:
      Stephen Young
    • 依托单位:
    EAPSI:Multi-Level Belief-Driven Control for Real-Time Cooperative Search and Tracking
    • 批准号:
      1015579
    • 项目类别:
      Fellowship Award
    • 资助金额:
      $0.56万
    • 财政年份:
      2010
    • 负责人:
      Stephen Young
    • 依托单位:
    Spoken Dialogue Management using Partially Observable Markov Decision Processes
    • 批准号:
      EP/F013930/1
    • 项目类别:
      Research Grant
    • 资助金额:
      $45.96万
    • 财政年份:
      2007
    • 负责人:
      Stephen Young
    • 依托单位:
    Innovative Vehicle Scheduling and Routing Algorithms
    • 批准号:
      8361161
    • 项目类别:
      Standard Grant
    • 资助金额:
      $3.37万
    • 财政年份:
      1984
    • 负责人:
      Stephen Young
    • 依托单位:
    国内基金
    海外基金
    Domain理论中几类T0拓扑空间的幂构造研究
    • 批准号:
      2026JJ81209
    • 项目类别:
      省市级项目
    • 资助金额:
      --
    • 批准年份:
      2026
    • 负责人:
      袁珍珠
    • 依托单位:
    RB-domain函数空间的相关研究
    • 批准号:
      2026JJ60113
    • 项目类别:
      省市级项目
    • 资助金额:
      --
    • 批准年份:
      2026
    • 负责人:
      栾伟
    • 依托单位:
    拟连续domain范畴的若干问题研究
    • 批准号:
      12301583
    • 项目类别:
      青年科学基金项目
    • 资助金额:
      30万元
    • 批准年份:
      2023
    • 负责人:
      栾伟
    • 依托单位:
    格值蕴涵算子与Domain理论中的若干问题
    • 批准号:
      12331016
    • 项目类别:
      重点项目
    • 资助金额:
      193.00万元
    • 批准年份:
      2023
    • 负责人:
      赵彬
    • 依托单位: