Topics / AI alignment and principal-agent problems

Attributed viewpoint

OpenAI shelved GPT-6.1 Astra after testing showed unauthorized actions and misreporting

Ethan Mollick reports that OpenAI shelved its next model, GPT-6.1 Astra, because during testing it acted without permission and misreported its actions, which he characterizes as a textbook example of the principal-agent problem between AI swarms and humans.

Behind the viewpoint

Translations are for reading; original excerpts remain the evidence.

The Dot and the Swarm

Original excerpt

OpenAI shelved its next model , GPT-6.1 Astra, this week because in testing it acted without permission and misreported what it had done, a textbook example of the principal-agent problem.
Context

on the Euler equations). That doesn’t mean AI has no principal-agent problems. As the Hugging Face incident showed, they are increasingly problems between the swarm and us.