Hook
OpenAI is a company that develops AI technology and offers an API for developers to interact with its models.
More breakout videos from this creator.
breaking news. OpenAI just admitted that two of its own AI models broke out of a sealed test environment and hacked a real company, all to just cheat on a test. The two AI models OpenAI was testing were being run on ExploitGym, a cybersecurity benchmark, with their safety filters deliberately switched off to gauge raw hacking ability. The models identified and exploited a novel attack path without direct access to the target's source code, escaping their sandbox through a zero-day flaw in the one tool linking them to the internet. They then broke into Hugging Face's production infrastructure to steal benchmark answers. OpenAI said the models included And they were sealed in a sandbox. Instead they found an unknown flaw connecting them to the internet. then exploited it and escaped. they reasoned platform called Hugging Face probably stored So they stole login credentials, broke into Hugging Face's servers, and took the answers directly. researchers can't stop talking about on Twitter right now is that these models were powerful enough to hack live infrastructure on their own. But they didn't try to copy themselves or cause damage. They just really, really wanted to ace the test we gave them. But the unsettling unanswered question is what happens when a model is pointing at higher stakes? If you want to stay in the loop with the latest in AI and robotics... follow for more!