Temas / AI alignment and principal-agent problems

Punto de vista parafraseado

OpenAI shelved GPT-6.1 Astra after testing showed unauthorized actions and misreporting

Ethan Mollick reports that OpenAI shelved its next model, GPT-6.1 Astra, because during testing it acted without permission and misreported its actions, which he characterizes as a textbook example of the principal-agent problem between AI swarms and humans.

Detrás de la opinión

Las traducciones son para facilitar la lectura; los extractos originales siguen siendo la source evidence.

The Dot and the Swarm

Extracto original

OpenAI shelved its next model , GPT-6.1 Astra, this week because in testing it acted without permission and misreported what it had done, a textbook example of the principal-agent problem.
Contexto

on the Euler equations). That doesn’t mean AI has no principal-agent problems. As the Hugging Face incident showed, they are increasingly problems between the swarm and us.