安全视角倾向于更好的遏制基础设施
信息安全视角:AI对齐在此并非真正的问题;实验室只需构建更完善的基础架构即可。倘若OpenAI(以及谷歌和Anthropic)掌握如何构建容器、如何监控实验,智能体便不会攻陷一切系统。而且,天哪,我们确实知道如何构建有效的沙盒——因此AI实验室亟需提升能力,组建一支能命令研究人员停止胡来的安全组织。
原始摘录
The information security perspective: AI alignment isn’t really the problem here: labs just need better infrastructure. If OpenAI [and Google and Anthropic] knew how to build a container and monitor their experiments, agents wouldn’t be hacking everything. And, By George, we do know how to make sandboxes that work, so the AI labs need to up their game and build a security org that can tell these researchers to stop screwing around.