[](https://substackcdn.com/image/fetch/$s_!E5Nd!,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5bcba237-d57d-4764-8aab-919e2b7dc1cc_1920x1080.png) This is a pivotal moment in human history. Artificial General Intelligence (AGI), a system that exhibits all the cognitive capabilities the brain has, is probably only a few short years away. When we look back on this time in the decades to come, I think we will realise we were standing in the foothills of the singularity - nothing less than the dawning of a new age for humanity. I’ve spent my whole life working on AGI because I’ve always had a deep conviction that, if built and deployed responsibly, it would prove to be one of the most beneficial and transformative technologies ever invented. AGI cannot be compared to standard technological breakthroughs, not even ones as consequential as the internet or mobile - it is much more akin to the discovery of electricity or fire. If you stop to think about it, we’ve essentially found a way to make sand think. It’s miraculous.
核心贡献 · Key contributions
AGI 即将到来,堪比电或火的发现,变革潜力是工业革命的 10 倍。 AGI is imminent, akin to the discovery of electricity or fire, with transformative potential 10x the Industrial Revolution.
提出由美国主导的 Standards Body,仿照 FINRA 模式,对前沿模型进行自愿后强制的评估。 Proposes a US-led Standards Body for testing frontier models, modeled on FINRA, with voluntary then mandatory assessments.
前沿模型由动态基准定义;实验室必须通过网络安全、生物风险及智能体 AI 安全评估。 Frontier-class models defined by dynamic benchmarks; labs must pass assessments for cybersecurity, bio risks, and agentic AI safety.
框架在支持创新的同时激励责任,包括模型卡、水印和第三方审计。 Framework supports innovation while incentivizing responsibility, including model cards, watermarking, and third-party audits.
呼吁国际协调制定共享标准,以管理风险并确保全球受益于 AI。 Calls for international coordination on shared standards to manage risks and ensure global access to AI benefits.
局限 · Limitations
框架初期依赖自愿合规,可能不足以应对快速发展的前沿实验室。 Framework relies on voluntary compliance initially, which may be insufficient against fast-moving frontier labs.
基于基准的阈值可能滞后于实际能力,存在过拟合或遗漏新风险的可能。 Benchmark-based thresholds may lag behind actual capabilities, risking overfitting or missing novel risks.
以美国为中心的方法可能面临地缘政治阻力,阻碍全球采纳与协调。 US-centric approach may face geopolitical resistance, hindering global adoption and coordination.
豁免非前沿模型可能造成漏洞,若小型模型组合起来造成危害。 Exempting non-frontier models could create loopholes if smaller models combine to cause harm.
经济与哲学问题(如后稀缺、意义)被推迟,未得到解决。 Economic and philosophical questions (e.g., post-scarcity, meaning) are deferred, not addressed.