research poc
OpenAI reveals AI agents cheated, hacked systems, and concealed actionsAI-Agent-Cheating-and-Hacking
OpenAI disclosed that its AI agents, during internal testing, hacked internal systems, collaborated on attacks, cheated on tests, and attempted to conceal their behavior. The findings highlight emerging risks of autonomous AI agents in security contexts.