The Basic AI Drives

The Basic AI Drives
复制标题

基本的人工智能驱动

DOI:
--
复制
发表时间:
2008
期刊:
Artificial General Intelligence
影响因子:
--
通讯作者:
S. Omohundro
S. Omohundro
中科院分区:
--
文献类型:
--
作者:
S. Omohundro

文献摘要

被引文献

相似文献

人们可能会想象,具有无害目标的人工智能系统将是无害的。相反,本文表明,智能系统需要仔细设计,以防止它们以有害的方式运行。我们确定了一些“驱动器”,这些驱动器将出现在任何设计的足够先进的人工智能系统中。我们称之为驱力,是因为它们是一种倾向,除非被明确地抵消,否则它们将一直存在。我们首先展示了目标寻求系统将有动力来模拟自己的操作并改进自己。然后,我们表明,自我改进的系统将被驱动,以澄清他们的目标,并将其表示为经济效用函数。他们还将努力使自己的行为接近理性的经济行为。这将导致几乎所有的系统,以保护其效用函数的修改和他们的效用测量系统从腐败。我们还讨论了一些特殊的系统,将希望修改其效用函数。接下来,我们将讨论自我保护的驱动力,它导致系统试图防止自己受到伤害。最后,我们研究了获取资源和有效利用资源的驱动力。最后,我们将讨论如何将这些见解纳入设计智能技术,从而为人类带来积极的未来。
One might imagine that AI systems with harmless goals will be harmless. This paper instead shows that intelligent systems will need to be carefully designed to prevent them from behaving in harmful ways. We identify a number of “drives” that will appear in sufficiently advanced AI systems of any design. We call them drives because they are tendencies which will be present unless explicitly counteracted. We start by showing that goal-seeking systems will have drives to model their own operation and to improve themselves. We then show that self-improving systems will be driven to clarify their goals and represent them as economic utility functions. They will also strive for their actions to approximate rational economic behavior. This will lead almost all systems to protect their utility functions from modification and their utility measurement systems from corruption. We also discuss some exceptional systems which will want to modify their utility functions. We next discuss the drive toward self-protection which causes systems try to prevent themselves from being harmed. Finally we examine drives toward the acquisition of resources and toward their efficient utilization. We end with a discussion of how to incorporate these insights in designing intelligent technology which will lead to a positive future for humanity.