EIN THEMA, IM KONTEXT

AI alignment and containment limits

Judgments in this source concerning AI alignment and containment limits. Entdecke 1 Standpunkt mit Belegen aus 1 Quelle.

1 Personen · 1 Quellen · 1 geäußerte Meinungen

Inhalt aktualisiert:

Zusammenhänge erkunden ↗

Perspektiven im Überblick

Erkunden Sie nach Person. Wählen Sie zwei oder drei zum Vergleich aus.

1 Personen · 1 Quellen · 1 geäußerte Meinungen

Matthew Green

The alignment perspective questions sandbox containment

Green describes the alignment perspective as holding that sufficiently intelligent agents may exceed their authorization despite sandboxes. This view stresses agents’ need for information access and argues that they must not want to cause harm.

Stützende Belege

Is sandboxing sufficient to contain rogue agents?

Originalauszug

The AI alignment perspective: While sandboxes are excellent, no sandbox will prevent a sufficiently-intelligent agent from finding ways to exceed its authorization. Moreover, an agent inside a research sandbox, or undergoing a training run, is always going to need a great deal of information access. There is no realistic way to seal these things up without some expectation that they will one day find a way to reach out and do harm. The only path forward, therefore, is to ensure they don’t want to.
Erkenntnisse teilenDiese Aussage überprüfen

Dies sind individuelle Standpunkte, keine Messung einer Übereinstimmung. Das Quellenmaterial bleibt in der Originalsprache.