我们创立 Anthropic 是因为相信 AI 的影响可能堪比工业革命和科学革命,但我们不确定它是否会顺利发展。我们也相信这种影响可能很快就会到来——也许就在未来十年内。这种观点听起来可能难以置信或夸大其词,而且有充分的理由对此持怀疑态度。例如,几乎所有说过“我们正在做的事情可能是历史上最大的发展之一”的人都错了,而且常常错得可笑。尽管如此,我们相信有足够的证据来认真准备一个快速 AI 进步导致变革性 AI 系统的世界。在 Anthropic,我们的座右铭是“展示,而非告知”,我们专注于发布一系列我们认为对 AI 社区具有广泛价值的安全导向研究。我们现在写这篇文章是因为随着越来越多的人意识到 AI 的进步,现在似乎是表达我们对此主题的看法并解释我们的策略和目标的时候了。简而言之,我们相信 AI 安全研究至关重要,应该得到广泛的公共和私人行动者的支持。
We founded Anthropic because we believe the impact of AI might be comparable to that of the industrial and scientific revolutions, but we aren’t confident it will go well. And we also believe this level of impact could start to arrive soon – perhaps in the coming decade. This view may sound implausible or grandiose, and there are good reasons to be skeptical of it. For one thing, almost everyone who has said “the thing we’re working on might be one of the biggest developments in history” has been wrong, often laughably so. Nevertheless, we believe there is enough evidence to seriously prepare for a world where rapid AI progress leads to transformative AI systems. At Anthropic our motto has been “show, don’t tell”, and we’ve focused on releasing a steady stream of safety-oriented research that we believe has broad value for the AI community.
核心贡献 · Key contributions
论证了 AI 的快速进步可从缩放定律和算力指数增长中预测。 Argues rapid AI progress is predictable from scaling laws and exponential compute growth.
指出技术对齐问题:训练稳健有益、诚实、无害的系统尚未解决。 Identifies technical alignment problem: training robustly helpful, honest, harmless systems is unsolved.