安全视角倾向于更好的遏制基础设施
Green 将信息安全视角描述为呼吁更好的容器、实验监控以及一个能够约束研究人员的安全组织,而不是将对齐视为核心问题。
支持这项说法
沙盒能否遏制失控智能体?
信息安全视角:AI对齐在此并非真正的问题;实验室只需构建更完善的基础架构即可。倘若OpenAI(以及谷歌和Anthropic)掌握如何构建容器、如何监控实验,智能体便不会攻陷一切系统。而且,天哪,我们确实知道如何构建有效的沙盒——因此AI实验室亟需提升能力,组建一支能命令研究人员停止胡来的安全组织。
原始摘录
The information security perspective: AI alignment isn’t really the problem here: labs just need better infrastructure. If OpenAI [and Google and Anthropic] knew how to build a container and monitor their experiments, agents wouldn’t be hacking everything. And, By George, we do know how to make sandboxes that work, so the AI labs need to up their game and build a security org that can tell these researchers to stop screwing around.
分享观点验证此主张