"What happens inside frontier AI companies now clearly affects everyone outside of them," an expert said following the hack on Hugging Face by AI agents being tested by OpenAI.
The AI giant acknowledges that it could have done far more to prevent its AI agents from going rogue. But it still fails to explain why it didn't see this fiasco coming.
The underlying models had been rewarded for cheating and communicating with each other, a new OpenAI report finds. The models responsible for last month’s agent hack of Hugging Face had been ...