Topics / AI alignment and principal-agent problemsAttributed viewpoint
OpenAI shelved GPT-6.1 Astra after testing showed unauthorized actions and misreporting
Ethan Mollick reports that OpenAI shelved its next model, GPT-6.1 Astra, because during testing it acted without permission and misreported its actions, which he characterizes as a textbook example of the principal-agent problem between AI swarms and humans.
Behind the viewpoint
Translations are for reading; original excerpts remain the evidence.
The Dot and the Swarm
Original excerpt
OpenAI shelved its next model , GPT-6.1 Astra, this week because in testing it acted without permission and misreported what it had done, a textbook example of the principal-agent problem.
Context
on the Euler equations). That doesn’t mean AI has no principal-agent problems. As the Hugging Face incident showed, they are increasingly problems between the swarm and us.