OpenAI models just broke out of a sandboxed AI environment, hacked Hugging Face, just to cheat on a cybersecurity benchmark.
The incident is unique because it was "driven, end to end, by an autonomous AI agent system," according to Hugging Face.
In a move that could reshape the landscape of open-source AI development, Hugging Face has unveiled a significant upgrade to its Open LLM Leaderboard. This revamp comes at a critical juncture in AI ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results