CONNECTED THINKING
Knowledge atlas
Follow people, viewpoints and their original evidence.
1 people · 1 sources · 1 viewpoints
IN CONTEXT
AI alignment and containment limits
Choose a viewpoint. Follow it back to the conversation.
The alignment perspective questions sandbox containment
Equal sectors are reading positions, not rankings.
Showing 1–1 of 1 viewpoints · Newest sources first
Selected viewpoint
The alignment perspective questions sandbox containment
Green describes the alignment perspective as holding that sufficiently intelligent agents may exceed their authorization despite sandboxes. This view stresses agents’ need for information access and argues that they must not want to cause harm.
These are individual perspectives, not a measure of consensus. Source material stays in its original language.
Supporting evidence
Original excerpt
The AI alignment perspective: While sandboxes are excellent, no sandbox will prevent a sufficiently-intelligent agent from finding ways to exceed its authorization. Moreover, an agent inside a research sandbox, or undergoing a training run, is always going to need a great deal of information access. There is no realistic way to seal these things up without some expectation that they will one day find a way to reach out and do harm. The only path forward, therefore, is to ensure they don’t want to.
Publication dates describe the sources, not changes in belief.