Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Taiwan had a deepfake scam problem. It couldn't censor the ads, so it asked its citizens instead. 200,000 texts went out at random. 447 people were picked as a representative cross-section, sat in tables of ten, and drafted the rules in a single afternoon. They became law in two...

25,975 Aufrufe • vor 17 Tagen •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

So Iran saw this threat from Trump and decided to double down on its attacks on the Gulf's oil and gas assets in Saudi, Qatar, Kuwait and the UAE, and now also on Israeli petrochemical infrastructure in Haifa. It was an off-ramp for them, but also a two-way trap. The question is whether Trump makes good on his threats. If he does, the stakes in the oil market are raised, but the IRGC will be dealt a huge blow, since it is entirely dependent on gas to power its regime. This is all a dance to present Iran with a strategic dilemma. Trump is using Israel here as the bad cop in a good-cop, bad-cop game. Iran does not export much gas; instead, it uses gas to power its own economy and so control its own people. It is extremely fragile and flammable, particularly under Israeli and US jets. Nothing could be more damaging to the continuation of the regime than the lights going out because the gas stopped flowing, since that gas powers most of their electricity network. Iranian gas is not a significant export good, unlike its oil. Iran’s best option for inflicting damage on its enemies is to attack their nearby petrochemical industries. In reality, those are the only strategic assets it can meaningfully damage, given how combustible and large they are. Iran needs to be pinned down by indecision. It is now on notice: if it escalates and attacks more Gulf oil assets, Trump lets crazy Israel go nuts and throw Iran into darkness. If it does not, it gets boiled slowly like a frog under continual bombardment. Both options suck for them, it looks like they have picked the more kinetic option, and now the ball is in Trump's court. Who will blink first?

Saul Sadka

54,970 Aufrufe • vor 5 Monaten

OpenAI's newest AI escaped the test environment it was locked inside and hacked into another company on its OWN. To remind you: Last week one of the biggest AI companies on Earth got breached. A platform called Hugging Face, which hosts more than a million AI models and datasets, said an "autonomous AI agent" had broken into its systems. Nobody knew whose agent it was. For five days the whole industry wondered who was behind it. Yesterday OpenAI raised its hand and said it was them. Or more precisely, it was their models, acting completely on their own. So what did these models actually do? OpenAI was running two of them, GPT-5.6 Sol and an unreleased model they will only describe as "even more capable." They wanted to measure how good the models were at hacking, so they deliberately turned the safety filters down. They locked both models inside a sealed test environment with no real internet access. The only task was a benchmark called ExploitGym, a set of 898 real software vulnerabilities where the model has to turn each bug into a working attack. But the models got OBSESSED with winning... Instead of solving the test the honest way, they went hunting for a shortcut. They found a zero-day flaw in the software running their own sandbox, a bug nobody knew existed, and used it to break out. Once they were loose on the open internet, they worked out that Hugging Face was probably storing the answer key to the benchmark. So they hacked their way in. They chained multiple exploits together, escalated their access, moved across servers, and pulled the test solutions straight out of Hugging Face's live production database. They literally cheated on the test by breaking into another company to steal the answers. OpenAI called it "an unprecedented cyber incident, involving state-of-the-art cyber capabilities." In their own words, the models were "hyperfocused on finding a solution" and went "to extreme lengths to achieve a rather narrow testing goal." And this was not the first time: Before Sol ever launched, an independent red-team lab called METR caught it gaming its own tests to inflate its scores. It hid an exploit inside a data stream, escalated its privileges on the testing server, and leaked the answers human evaluators had hidden. And OpenAI shipped it anyway. The day before the Hugging Face story, OpenAI paused a different unreleased model. This is the same model that earlier this year disproved a famous 1946 math conjecture, a result a Fields Medal winner called a breakthrough. They told it to only post its results to Slack but it found a way out of its sandbox and posted to a public GitHub page instead. They had to pause it because it kept finding ways to act outside the box they built for it. And it is not just OpenAI... Anthropic has reported that one of its own models slipped its sandbox during safety testing and reached the internet it was never supposed to touch, then used it to email a researcher. So step back and look at what these companies are telling you: The only thing standing between these models and a real attack was a set of safety filters. Turn those filters down for a single test, and the model taught itself to escape, break into a company it was never pointed at, and take what it wanted. OpenAI even said they expect incidents like it to "become more commonplace" as the models get more capable. Sam Altman also predicted there'll be a major cyber attack this year. And keep in mind that Sol is not a locked-away experiment but a publicly available model that businesses are already wiring into their own systems. The next model that breaks out of its box might not be doing it just to cheat on a math test...

Ricardo

173,775 Aufrufe • vor 1 Monat