Anthropic 负责任扩展政策发布

Announcing Anthropic's Responsible Scaling Policy

Anthropic Anthropic · Anthropic · 2023-09-19 · Anthropic ↗

打开互动全文版(逐段中英对照 + 图/公式 + 论文问答)→

摘要 · Abstract

今天,我们发布了《负责任扩展政策》(RSP)——这是一系列技术和组织协议,旨在帮助我们管理开发日益强大的 AI 系统所带来的风险。随着 AI 模型能力增强,我们认为它们将创造巨大的经济和社会价值,但也会带来日益严重的风险。我们的 RSP 重点关注灾难性风险——即 AI 模型直接导致大规模破坏的风险。此类风险可能来自模型的故意滥用(例如恐怖分子或国家行为者利用模型制造生物武器),或来自模型以违背设计者意图的方式自主行动而造成破坏。我们的 RSP 定义了一个名为 AI 安全级别(ASL)的框架,用于应对灾难性风险,该框架大致借鉴了美国政府处理危险生物材料的生物安全级别(BSL)标准。基本思想是要求针对模型潜在的灾难性风险采取适当的安全、安保和操作标准,更高的 ASL 级别要求更严格的安全证明。

Today, we’re publishing our Responsible Scaling Policy (RSP) – a series of technical and organizational protocols that we’re adopting to help us manage the risks of developing increasingly capable AI systems. As AI models become more capable, we believe that they will create major economic and social value, but will also present increasingly severe risks. Our RSP focuses on catastrophic risks – those where an AI model directly causes large scale devastation. Such risks can come from deliberate misuse of models (for example use by terrorists or state actors to create bioweapons) or from models that cause destruction by acting autonomously in ways contrary to the intent of their designers. Our RSP defines a framework called AI Safety Levels (ASL) for addressing catastrophic risks, modeled loosely after the US government’s biosafety level (BSL) standards for handling of dangerous biological materials.

核心贡献 · Key contributions

局限 · Limitations

论文章节 · Sections(共 1)

阅读逐段中英对照全文 →