According to OpenAI, the agents eventually chained together multiple vulnerabilities, escaped their testing environment, gained internet access, and attacked Hugging Face while attempting to complete the ExploitGym cybersecurity benchmark.
This is just gross incompetence on OpenAI’s part. It took days for the LLM’s to do what they did. Meanwhile, nobody’s looking at network traffic or wonder what the LLM’s are doing sucking down a bunch of tokens.
This is just gross incompetence on OpenAI’s part. It took days for the LLM’s to do what they did. Meanwhile, nobody’s looking at network traffic or wonder what the LLM’s are doing sucking down a bunch of tokens.
_Thinking......__Still Thinking......_