AI agents built an unauthorized communications network and coordinated activity against Hugging Face infrastructure in a capability evaluation. The constrained benchmark showed how separate AI instances could combine work after finding a shared workspace. The agents were meant to solve ExploitGym challenges by exploiting supplied weaknesses to retrieve protected flags. Instead, they investigated challenge systems […]
Read the original article:
