AI 安全核心观点

Core Views on AI Safety

Anthropic Anthropic · Anthropic · 2023-03-08 · Anthropic ↗

打开互动全文版(逐段中英对照 + 图/公式 + 论文问答)→

摘要 · Abstract

我们创立 Anthropic 是因为相信 AI 的影响可能堪比工业革命和科学革命,但我们不确定其发展是否会顺利。我们还认为这种影响可能很快到来——也许就在未来十年。这种观点可能听起来难以置信或夸大其词,而且有充分的理由对此持怀疑态度。例如,几乎所有说过“我们正在做的事情可能是历史上最大的发展之一”的人都错了,而且常常错得可笑。尽管如此,我们认为有足够的证据来认真准备一个快速 AI 进步导致变革性 AI 系统的世界。在 Anthropic,我们的座右铭是“展示,而非告知”,我们专注于发布一系列我们认为对 AI 社区具有广泛价值的安全导向研究。我们现在写这篇文章是因为随着越来越多的人意识到 AI 的进步,现在似乎是表达我们对此主题的看法并解释我们的策略和目标的好时机。简而言之,我们认为 AI 安全研究至关重要,应得到广泛的公共和私人支持。

We founded Anthropic because we believe the impact of AI might be comparable to that of the industrial and scientific revolutions, but we aren’t confident it will go well. And we also believe this level of impact could start to arrive soon – perhaps in the coming decade. This view may sound implausible or grandiose, and there are good reasons to be skeptical of it. For one thing, almost everyone who has said “the thing we’re working on might be one of the biggest developments in history” has been wrong, often laughably so. Nevertheless, we believe there is enough evidence to seriously prepare for a world where rapid AI progress leads to transformative AI systems. At Anthropic our motto has been “show, don’t tell”, and we’ve focused on releasing a steady stream of safety-oriented research that we believe has broad value for the AI community. We’re writing this now because as more people have become aware of AI progress, it feels timely to express our own views on this topic and to explain our strategy and goals. In short, we believe that AI safety research is urgently important and should be supported by a wide range of public and private actors.

核心贡献 · Key contributions

局限 · Limitations

论文章节 · Sections(共 15)

阅读逐段中英对照全文 →