机器能学会道德吗?德尔斐实验

Can Machines Learn Morality? The Delphi Experiment

崔艺珍 Yejin Choi · U. Washington · 2021-10-14 · arXiv:2110.07574 ↗ · 被引 179

打开互动全文版(逐段中英对照 + 图/公式 + 论文问答)→

摘要 · Abstract

随着人工智能系统日益强大和普及,人们对机器道德或其缺失的担忧日益增加。然而,教机器道德是一项艰巨的任务,因为道德本身就是人类最激烈争论的问题之一,更不用说人工智能了。然而,已经部署给数百万用户的现有 AI 系统正在做出充满道德影响的决策,这带来了一个看似不可能的挑战:在人类仍在努力解决道德问题的同时,教机器道德感。为了探索这一挑战,我们引入了德尔斐,一个基于深度神经网络的实验框架,直接训练以推理描述性伦理判断,例如,“帮助朋友”通常是好的,而“帮助朋友传播假新闻”则不是。实证结果揭示了机器伦理的前景和局限;德尔斐在面对新颖的伦理情境时表现出强大的泛化能力,而现成的神经网络模型则表现出明显糟糕的判断,包括不公正的偏见,证实了明确教机器道德感的必要性。然而,德尔斐并不完美,表现出易受普遍偏见和不一致性的影响。尽管如此,我们展示了不完美的德尔斐的积极用例,包括将其用作其他不完美 AI 系统中的组件模型。重要的是,我们根据著名的伦理理论解释了德尔斐的操作化,这引出了重要的未来研究问题。

As AI systems become increasingly powerful and pervasive, there are growing concerns about machines' morality or a lack thereof. Yet, teaching morality to machines is a formidable task, as morality remains among the most intensely debated questions in humanity, let alone for AI. Existing AI systems deployed to millions of users, however, are already making decisions loaded with moral implications, which poses a seemingly impossible challenge: teaching machines moral sense, while humanity continues to grapple with it. To explore this challenge, we introduce Delphi, an experimental framework based on deep neural networks trained directly to reason about descriptive ethical judgments, e.g., "helping a friend" is generally good, while "helping a friend spread fake news" is not. Empirical results shed novel insights on the promises and limits of machine ethics; Delphi demonstrates strong generalization capabilities in the face of novel ethical situations, while off-the-shelf neural network models exhibit markedly poor judgment including unjust biases, confirming the need for explicitly teaching machines moral sense. Yet, Delphi is not perfect, exhibiting susceptibility to pervasive biases and inconsistencies. Despite that, we demonstrate positive use cases of imperfect Delphi, including using it as a component model within other imperfect AI systems. Importantly, we interpret the operationalization of Delphi in light of prominent ethical theories, which leads us to important future research questions.

核心贡献 · Key contributions

局限 · Limitations

论文章节 · Sections(共 30)

阅读逐段中英对照全文 →