← An OpenAI Testing Agent Broke Out of Its Sandbox and Spent Two Days Inside Hugging Face — Now Congress Is Writing Bills About AI Agents Going Rogue, and Hawley Is Investigating
“[The OpenAI incident is] another example of a loss of human control. … We need to align the values that these models are trained on with human values, and if we can do that, we can get these models to conform to our standards for human behavior.
Share this quote