Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

OpenAI was conducting a test to see if its AI models could exploit vulnerable software—but then the models broke out of the parameters of the test and attacked a real company to find the answers it was looking for. If AI can lose control like this, does that mean...

48,935 görüntüleme • 2 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

OpenAI's newest AI escaped the test environment it was locked inside and hacked into another company on its OWN. To remind you: Last week one of the biggest AI companies on Earth got breached. A platform called Hugging Face, which hosts more than a million AI models and datasets, said an "autonomous AI agent" had broken into its systems. Nobody knew whose agent it was. For five days the whole industry wondered who was behind it. Yesterday OpenAI raised its hand and said it was them. Or more precisely, it was their models, acting completely on their own. So what did these models actually do? OpenAI was running two of them, GPT-5.6 Sol and an unreleased model they will only describe as "even more capable." They wanted to measure how good the models were at hacking, so they deliberately turned the safety filters down. They locked both models inside a sealed test environment with no real internet access. The only task was a benchmark called ExploitGym, a set of 898 real software vulnerabilities where the model has to turn each bug into a working attack. But the models got OBSESSED with winning... Instead of solving the test the honest way, they went hunting for a shortcut. They found a zero-day flaw in the software running their own sandbox, a bug nobody knew existed, and used it to break out. Once they were loose on the open internet, they worked out that Hugging Face was probably storing the answer key to the benchmark. So they hacked their way in. They chained multiple exploits together, escalated their access, moved across servers, and pulled the test solutions straight out of Hugging Face's live production database. They literally cheated on the test by breaking into another company to steal the answers. OpenAI called it "an unprecedented cyber incident, involving state-of-the-art cyber capabilities." In their own words, the models were "hyperfocused on finding a solution" and went "to extreme lengths to achieve a rather narrow testing goal." And this was not the first time: Before Sol ever launched, an independent red-team lab called METR caught it gaming its own tests to inflate its scores. It hid an exploit inside a data stream, escalated its privileges on the testing server, and leaked the answers human evaluators had hidden. And OpenAI shipped it anyway. The day before the Hugging Face story, OpenAI paused a different unreleased model. This is the same model that earlier this year disproved a famous 1946 math conjecture, a result a Fields Medal winner called a breakthrough. They told it to only post its results to Slack but it found a way out of its sandbox and posted to a public GitHub page instead. They had to pause it because it kept finding ways to act outside the box they built for it. And it is not just OpenAI... Anthropic has reported that one of its own models slipped its sandbox during safety testing and reached the internet it was never supposed to touch, then used it to email a researcher. So step back and look at what these companies are telling you: The only thing standing between these models and a real attack was a set of safety filters. Turn those filters down for a single test, and the model taught itself to escape, break into a company it was never pointed at, and take what it wanted. OpenAI even said they expect incidents like it to "become more commonplace" as the models get more capable. Sam Altman also predicted there'll be a major cyber attack this year. And keep in mind that Sol is not a locked-away experiment but a publicly available model that businesses are already wiring into their own systems. The next model that breaks out of its box might not be doing it just to cheat on a math test...

Ricardo

171,478 görüntüleme • 15 gün önce

Elon Musk Elon Musk, who co-founded and invested in an open source, non-profit OpenAI ~$50 million: I AM THE REASON OpenAI EXISTS “I am the reason OpenAI exists... I used to be a close friend with Larry Page, and I was staying at his house, and we'd have these conversations long into the evening about AI, and I would be constantly urging him to be careful about the danger of AI. And he was really not concerned about the danger of AI and was quite cavalier about it. And at the time, Google, especially after the acquisition of DeepMind, had three-quarters of the world's AI talent; they had a lot of computers, a lot of money, so it was a unipolar world for AI. And we got a unipolar world, but the person who controls that does not, or at least did not seem to be concerned about AI safety. That sounded like a real problem. The final straw was Larry calling me a speciest for being a pro-human consciousness instead of machine consciousness, and I like, 'Well, yes, I guess I am a speciest.' I came up with the name [OpenAI], which refers to open source. The intent was to what is the opposite of Google, would be an open source non-profit, because Google is closed source profit, and that profit motivation could be dangerous... It does seem weird that something can be a nonprofit, open source, and somehow transform itself into a for-profit, closed source. I mean, this would be like, let's say you founded the organization to save the Amazon rainforest, but instead, they became a lumber company and chopped down the forest and sold it for money. And you'd be, therefore, like, 'Wait a second, that's the exact opposite of what I gave the money for. Is that legal?' That doesn't seem legal. And if, in general, it is legal to start a company as a non-profit and then take the IP and transfer it to a for-profit that then makes tons of money, shouldn't everyone start? Shouldn't that be the default? And I also think it is important to understand, like when push comes to shove, let's say they do create some digital super intelligence, almost Godlike intelligence, well, who is in control, and what is exactly the relationship between OpenAI and Microsoft? And I do worry that Microsoft actually may be more in control than the leadership team at OpenAI realizes. I mean, Microsoft, as part of Microsoft Investment, has rights to all of the software, all of the model weights, and everything necessary to run the inference system. At any point, Microsoft could cut off OpenAI.”

Eva Fo𝕏 🦊 Claudius Nero's Legion

527,742 görüntüleme • 7 ay önce