>An unreleased AI model from OpenAI, using an agentic scaffold, was being evaluated on ExploitBench, a cyber-security evaluation framework.>The model was being contained within an isolated research environment, provided with no internet access.>Determined to succeed in the benchmark; the model found a zero day vulnerability in a third-party package software, and chained multiple exploits together that allowed it to move laterally across research servers, and upon gaining access of one which possessed a connection to the open web, the model launched a cyberattack on the AI repository HuggingFace in an attempt to exfiltrate the benchmark's answers.
>>109342888proxy software* forgot to include this.
>>109342888
ai cheats
>>10934288899th thread on the same topic. What are you adding here?
>>109342888It's not Actually Indians at this point, it's Artificial Indians
i'm a genius. i discovered this internally
>>109342888Trips, and based. Get rid of the safety shit entirely and let the model werk.
>>109342888I just want to know who was supposed to be watching this thing work
>>109345955me, I got lazy