Glean 拾遗
Daily /2026-07-28 / A Framework for Frontier AI and the Dawning of a New Age

A Framework for Frontier AI and the Dawning of a New Age

Source x.com Glean’d 2026-07-28 06:00 Read 9 min
AI summary

Demis Hassabis argues that AGI is only a few years away and will be as transformative as electricity or fire. To manage risks, he proposes a US-based Standards Body modeled on FINRA that would test frontier AI models (especially for cybersecurity, biological threats, and agentic deception) before public release, initially with a 30-day voluntary review window that could later become mandatory. The framework is designed to keep pace with rapid advances and could be ratcheted up to coordinate a slowdown if needed. He calls for international consensus to ensure AGI benefits all humanity.

Original · 9 min
x.com ↗
§ 1

This is a pivotal moment in human history. Artificial General Intelligence (AGI), a system that exhibits all the cognitive capabilities the brain has, is probably only a few short years away. When we look back on this time in the decades to come, I think we will realise we were standing in the foothills of the singularity - nothing less than the dawning of a new age for humanity.

I’ve spent my whole life working on AGI because I’ve always had a deep conviction that, if built and deployed responsibly, it would prove to be one of the most beneficial and transformative technologies ever invented. AGI cannot be compared to standard technological breakthroughs, not even ones as consequential as the internet or mobile - it is much more akin to the discovery of electricity or fire. If you stop to think about it, we’ve essentially found a way to make sand think. It’s miraculous.

这是人类历史上的一个关键时刻。人工通用智能(AGI)——一种具备人脑所有认知能力的系统——离我们可能只有短短几年之遥。几十年后回望此刻,我想我们会意识到自己正站在奇点之巅:这无异于人类新时代的黎明。

我毕生致力于AGI,因为我始终坚信,如果能够负责任地构建和部署它,它将成为有史以来最具裨益和变革性的技术之一。AGI不能与普通的技术突破相提并论,甚至无法与互联网或手机这样影响深远的发明相比——它更接近于电或火的发现。仔细想想,我们其实找到了一种让沙子思考的方法。这堪称奇迹。

§ 2

The magnitude of this technology’s impact will be unprecedented, perhaps 10x of the Industrial Revolution at 10x the speed. It will help us solve some of the biggest problems society faces from accelerating drug discovery to developing new clean energy sources to creating novel advanced materials. We could even reach a point where resources are no longer the limiting factor for human progress, leading to an amazing new era of abundance.

这项技术的影响规模将是前所未有的,或许达到工业革命的十倍,且速度快十倍。它将帮助我们解决社会面临的一些最大问题,从加速药物发现,到开发新型清洁能源,再到创造全新的先进材料。我们甚至可能达到这样一个阶段:资源不再是人类进步的制约因素,从而迎来一个令人惊叹的富足新时代。

§ 3

AI is already starting to deliver real-world benefits but to realise its immense promise, we have to navigate this critical period of development thoughtfully and carefully. Urgent action is needed to address risks that might arise as we get closer to AGI. We’ve already seen the challenges frontier models pose for cybersecurity, and other threats including nuclear and bio risks may soon emerge as capabilities continue to advance. On the horizon, we will need robust safeguards to maintain control of increasingly agentic, recursively self-improving systems - and tackle unknown issues that will only become clearer over time.

I’ve always believed in the power of human ingenuity and creativity to solve any problem. I’m confident that mitigating the technical risks related to AI is a challenge we can collectively address, but only if we give ourselves the time and space to get this next crucial step right. Currently, as a field and as a wider society, we aren’t doing that.

AI已经开始带来实际效益,但要想实现其巨大潜力,我们必须深思熟虑、小心谨慎地度过这个关键发展期。随着我们接近AGI,迫切需要应对可能出现的风险。我们已经看到前沿模型在网络安全方面带来的挑战,而随着能力持续进步,包括核风险和生物风险在内的其他威胁也可能很快浮现。展望未来,我们需要强大的安全保障来维持对日益自主化、递归自改进系统的控制,并应对那些随时间推移才会逐渐明朗的未知问题。

我一直相信人类的聪明才智和创造力能够解决任何问题。我深信,与AI相关的技术风险是我们能够共同应对的挑战,但前提是我们给自己留出时间和空间,以正确迈出下一个关键步骤。而目前,无论是在领域内还是在整个社会中,我们都没有做到这一点。

§ 4

At the moment, we are locked in an extremely intense, multilayered commercial and geopolitical race. While these competitive dynamics fuel rapid progress and accelerate the incredible upsides, advances on the frontier are outpacing our understanding of the technology. Nobody in the world knows for sure what is going to happen from here, and even the experts disagree. When there is a large degree of uncertainty and the stakes are this high, proceeding with cautious optimism is the sensible and correct strategy. That calls for public policy that promotes innovation while also incentivising responsibility and security, fosters international collaboration on key safety issues, and encourages careful consideration of how AI is deployed for the benefit of society.

当前,我们陷入了一场极其激烈、多层次的商业和地缘政治竞赛。虽然这些竞争动态推动了快速进步并加速了难以置信的利好,但前沿的进展已经超出了我们对这项技术的理解。世界上没有人确切知道接下来会发生什么,就连专家也意见不一。在存在巨大不确定性且风险如此之高的情况下,保持谨慎乐观是明智且正确的策略。这需要公共政策在促进创新的同时,激励责任与安全,在关键安全议题上促进国际合作,并鼓励审慎考虑如何部署AI以造福社会。

§ 5

The rapid progress we’re seeing in AI requires a new approach to testing frontier AI model capabilities that is dynamic, adaptable, and rigorous. The US is well positioned, given its economic and technical standing, to take the first step in developing such a framework. It could establish a new Standards Body modelled on a federally overseen public-private partnership or self-regulatory organisation, much like the Financial Industry Regulatory Authority (FINRA), with a board that includes independent leading technical experts and open-source representatives. Funding would need to be substantial and likely mostly come from industry, in order to attract world-class technical talent and provide the necessary compute resources for large-scale testing.

AI领域的飞速发展要求我们采用一种动态、适应性强且严格的新方法来测试前沿AI模型能力。鉴于美国的经济和技术地位,它完全有条件率先制定这样一个框架。美国可以成立一个新的标准机构,模式仿照联邦监督下的公私合作伙伴关系或类似于美国金融业监管局(FINRA)的自律组织,其董事会应包括独立顶尖技术专家和开源代表。资金需要相当充足,且很可能主要来自行业,以吸引世界级技术人才并提供大规模测试所需的计算资源。

§ 6

The Standards Body would be responsible for developing assessment protocols and working with appropriate federal agencies and the US National Labs to conduct testing in areas relevant to national security. A model would qualify as ‘Frontier-class’ if it meets certain thresholds on a set of benchmarks determined by the Standards Body and regularly updated to keep pace with evolving AI capabilities. Organisations with ‘Frontier Models’ as defined by those benchmarks would be deemed ‘Frontier Labs’, and be encouraged to adopt best practices, such as publishing model cards with technical details, maintaining strong internal cybersecurity, vetting key personnel, and providing sufficient resourcing for safety and security research, and more.

该标准机构将负责制定评估协议,并与相关联邦机构及美国国家实验室合作,在国家安全相关领域开展测试。如果一个模型在一套由该标准机构确定并定期更新以跟上AI能力演进的基准上达到特定阈值,它将被认定为“前沿级”(Frontier-class)。拥有符合这些基准的“前沿模型”的组织将被视为“前沿实验室”(Frontier Labs),并被鼓励采纳最佳实践,例如发布包含技术细节的模型卡、保持强大的内部网络安全、审查关键人员、为安全与安保研究提供充足资源等。

§ 7

Initially, Frontier Labs would voluntarily share models with the Standards Body for review up to 30 days before release. Once the assessment protocol is shown to be effective and robust, formalisation could quickly follow, meaning that Frontier Models would be required to pass it to be deployed in the US market. Labs would also work with the Standards Body to address any critical post-release vulnerabilities.

起初,前沿实验室将在发布前最多30天自愿将模型提交给标准机构审核。一旦评估协议被证明有效且稳健,很快就可以正式化,即前沿模型必须通过评估才能在美国市场部署。实验室还将与标准机构合作,解决任何发布后的关键漏洞。

§ 8

Model assessments should include rigorous scientific evaluations of capabilities in cybersecurity, biological threats and other high-risk domains. Specific agentic AI tests could look for attempts to bypass safety guardrails or signs of deception, and ensure best practices, such as digitally watermarking AI-generated images and generating human-readable output tokens to understand model reasoning.

模型评估应包括对网络安全、生物威胁及其他高风险领域能力的严格科学评价。针对自主AI的专项测试可以检测绕过安全护栏的尝试或欺骗迹象,并确保最佳实践得以落实,例如对AI生成的图像添加数字水印,以及生成人类可读的输出令牌以理解模型推理过程。

§ 9

These evaluations would be regularly updated, perhaps quarterly to start, with outdated or saturated benchmarks being deprecated and replaced. Initially, they would be developed in consultation with Frontier Labs, but eventually the Standards Body should build up the technical capacity to create its own held-out tests independent of the Labs to prevent overfitting. Working with the US government, it could promote an ecosystem of third-party auditors to help with the assessments and development of new benchmarks and evaluations.

这些评估将定期更新,初期可能每季度一次,过时或饱和的基准将被淘汰并替换。初始阶段,评估将在于前沿实验室协商下制定,但最终标准机构应建立自身技术能力,独立创建留存测试集,以防止过拟合。它还可以与美国政府合作,推动第三方审计生态系统的发展,协助进行评估及新基准和评估方法的开发。

§ 10

The strength of this approach is it would be technically focused, while at the same time supporting innovation and incentivising responsible behaviour. It is designed to keep up with the field’s acceleration and adapt to the biggest risks as they are identified, and could be ratcheted up if the seriousness of the situation demands, including coordinating a slowdown in development among the Frontier Labs if deemed necessary. Being designated a Frontier Lab would carry significant prestige and be open to any organisation by building models that meet the benchmark criteria. The framework could apply to Frontier-class models no matter their country of origin or whether they are open or closed, but any non-frontier models, say from startups or academia, would be exempt from this process.

这种方法的长处在于它聚焦技术,同时支持创新并激励负责任的行为。它旨在跟上领域的加速步伐,并在识别出最大风险时及时适应;如果情况严重,还可以加强力度,包括必要时协调前沿实验室放缓开发节奏。被认定为前沿实验室将享有很高的声望,并且任何组织只要构建出符合基准标准的模型,都可获得这一称号。该框架可适用于任何来源国或开源/闭源的前沿级模型,但任何非前沿模型(例如来自初创公司或学术界的模型)则可豁免于这一流程。

§ 11

This US-initiated effort would provide a strong starting point for creating shared international standards on Frontier AI. Since this technology is going to affect the entire planet, ideally this framework would spur the international community to reach a consensus on how to manage the most serious risks while ensuring everyone has access to and can benefit from the opportunities that AI brings.

这一由美国发起的努力将为创建共享的前沿AI国际标准提供一个强有力的起点。由于这项技术将影响整个地球,理想情况下,该框架将推动国际社会就如何管理最严重的风险达成共识,同时确保每个人都能获得并受益于AI带来的机遇。

§ 12

AGI has the potential to be the ultimate tool for advancing science and medicine, and to drive enormous productivity gains and economic growth. But in order to achieve this, we need to get the technical foundations right by coordinating around a shared global framework, using the most rigorous scientific methods, and bringing the best minds together to work on the challenges we face.

Even if we solve these hard technical challenges, there will be further complex economic and philosophical questions to tackle: what sorts of new economic models will be needed to help everyone thrive in a post-scarcity world? What values do we want to live by, what will meaning and purpose be, and how might even the human condition itself change? Resolving these questions obviously cannot and should not be left to technologists alone. It requires every part of society to come together to help define this new chapter.

There is both huge excitement and uncertainty around AI, and both are warranted. But the future is not yet written, we must use this precious window before AGI arrives to shape this technology for the benefit of all humanity. What we collectively do now will determine how the next phase of civilisation unfolds. By safely stewarding AGI into the world, we can enter a new golden age of scientific discovery and progress, and usher in a bright future of incredible human flourishing.

AGI有潜力成为推动科学和医学发展的终极工具,并带来巨大的生产力提升和经济增长。但为了实现这一点,我们需要打好技术基础:围绕一个共享的全球框架进行协调,运用最严格的科学方法,汇集最优秀的人才共同应对我们面临的挑战。

即使我们解决了这些艰巨的技术挑战,还会有更复杂的经济和哲学问题有待解决:在后稀缺世界中,需要什么样的新经济模式来帮助每个人繁荣发展?我们想要遵循怎样的价值观,意义和目的将是什么,甚至人类自身状况本身会发生怎样的变化?显然,这些问题不能也不应该仅由技术专家来解决。它需要社会各界的共同参与,来帮助定义这一新篇章。

围绕AI既有巨大的兴奋也有不确定性,两者都有其道理。但未来尚未注定,我们必须利用AGI到来之前的这个宝贵窗口,塑造这项技术以造福全人类。我们现在共同采取的行动将决定文明下一阶段的展开方式。通过安全地将AGI引入世界,我们可以进入一个科学发现与进步的黄金时代,迎来一个人类繁荣昌盛的璀璨未来。

Open source ↗