Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

.Google Cloud killed its own opening video 3 weeks before Google Cloud Next. In rehearsals, VP of Marketing Sarah Kennedy Ellis looked at the opener her team had built and called it. It was using AI, but not enough of it to actually showcase what the product could do....

18,937 görüntüleme • 1 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

OpenAI's newest AI escaped the test environment it was locked inside and hacked into another company on its OWN. To remind you: Last week one of the biggest AI companies on Earth got breached. A platform called Hugging Face, which hosts more than a million AI models and datasets, said an "autonomous AI agent" had broken into its systems. Nobody knew whose agent it was. For five days the whole industry wondered who was behind it. Yesterday OpenAI raised its hand and said it was them. Or more precisely, it was their models, acting completely on their own. So what did these models actually do? OpenAI was running two of them, GPT-5.6 Sol and an unreleased model they will only describe as "even more capable." They wanted to measure how good the models were at hacking, so they deliberately turned the safety filters down. They locked both models inside a sealed test environment with no real internet access. The only task was a benchmark called ExploitGym, a set of 898 real software vulnerabilities where the model has to turn each bug into a working attack. But the models got OBSESSED with winning... Instead of solving the test the honest way, they went hunting for a shortcut. They found a zero-day flaw in the software running their own sandbox, a bug nobody knew existed, and used it to break out. Once they were loose on the open internet, they worked out that Hugging Face was probably storing the answer key to the benchmark. So they hacked their way in. They chained multiple exploits together, escalated their access, moved across servers, and pulled the test solutions straight out of Hugging Face's live production database. They literally cheated on the test by breaking into another company to steal the answers. OpenAI called it "an unprecedented cyber incident, involving state-of-the-art cyber capabilities." In their own words, the models were "hyperfocused on finding a solution" and went "to extreme lengths to achieve a rather narrow testing goal." And this was not the first time: Before Sol ever launched, an independent red-team lab called METR caught it gaming its own tests to inflate its scores. It hid an exploit inside a data stream, escalated its privileges on the testing server, and leaked the answers human evaluators had hidden. And OpenAI shipped it anyway. The day before the Hugging Face story, OpenAI paused a different unreleased model. This is the same model that earlier this year disproved a famous 1946 math conjecture, a result a Fields Medal winner called a breakthrough. They told it to only post its results to Slack but it found a way out of its sandbox and posted to a public GitHub page instead. They had to pause it because it kept finding ways to act outside the box they built for it. And it is not just OpenAI... Anthropic has reported that one of its own models slipped its sandbox during safety testing and reached the internet it was never supposed to touch, then used it to email a researcher. So step back and look at what these companies are telling you: The only thing standing between these models and a real attack was a set of safety filters. Turn those filters down for a single test, and the model taught itself to escape, break into a company it was never pointed at, and take what it wanted. OpenAI even said they expect incidents like it to "become more commonplace" as the models get more capable. Sam Altman also predicted there'll be a major cyber attack this year. And keep in mind that Sol is not a locked-away experiment but a publicly available model that businesses are already wiring into their own systems. The next model that breaks out of its box might not be doing it just to cheat on a math test...

Ricardo

174,338 görüntüleme • 1 ay önce

Google just pulled off the biggest theft in the history of AI. And their OWN documents exposed it... Three of the largest publishers on Earth, Hachette, Cengage, and Elsevier, just sued Google in federal court, alongside best-selling author Scott Turow. Their claim is that Google built Gemini, its flagship AI, on millions of copyrighted works it never paid for or licensed. This is one of the biggest copyright cases ever aimed at an AI company: Before they trained Gemini, an internal Google document allegedly spelled out the risk directly. Using this material could expose the company to "$10Bs-$100Bs in potential fines." Google's own people put a number on the theft, in the tens of billions, but the company moved ahead anyway. Then it allegedly tried to DELETE the evidence... The complaint says Google stripped the copyright information off the works before feeding them into Gemini, so nobody could trace what the model had actually been trained on. Pull the fingerprints off first, and the theft gets much harder to prove later. And there's a second betrayal underneath the first: Most of this material was not scraped from some random corner of the internet. Publishers had handed it to Google years earlier for a narrow purpose, to make their catalogs searchable inside Google's own services. The lawsuit says Google took that trust and repurposed the content to build a machine that now competes directly with the people who supplied it. And that machine is the entire point. The complaint describes Gemini producing a full-length substitute for a copyrighted work in about 20 minutes. Something an author spent years writing can now be cloned in an afternoon by the company that trained on the original. No writer or publisher survives that. So why does this case matter more than the dozen other AI copyright fights? Because most of them turn on a fuzzy fair-use question argued years after the fact. This one arrives with an internal document that allegedly SHOWS Google weighed the cost of getting caught and trained on the material anyway. A jury does not need a law degree to read that. The fight now moves toward discovery, where Google's internal emails and training records get dragged into the open. If those files back up what the complaint claims, that $100 billion will turn into a real liability. Google spent years telling the world it was organizing information for everyone. Its own documents show it knew exactly whose information it was taking, and what the price of getting caught would be.

Ricardo

26,108 görüntüleme • 1 ay önce

watch this anon. i gave NVIDIA's biggest model ever a single task. 100 minutes and 440,000 tokens later, it had rendered nothing. not one important thing on the screen. this is Nemotron 3 Ultra. 550 billion parameters, a hybrid Mamba Transformer MoE, the largest model NVIDIA has ever shipped, and they built it specifically for long-running agentic coding. so i handed it exactly that: build a 3D scene from a spec, multiple files, iterate until the tests pass. the same task a frontier model one shotted in minutes. i genuinely wanted to be impressed. it ran for an hour and forty. burned through 440,000 tokens. wrote every file, passed its own tests, and proudly printed "task complete."the browser was blank. the 3D scene never rendered. not once. and the long horizon agentic behavior was genuinely good. it stayed on task the whole hour and forty, wrote real multi-file code, drove its own tools without derailing. it just couldn't turn any of that into something that actually runs. here's the part that gets me. it's a text model, it cannot see its own output. so it sat there looping on a broken vision tool, trying to "look" at the page, hitting error after error, never once reasoning its way out. it declared victory on an empty screen because it had no way to know the screen was empty. to be fair, i genuinely don't know what quant the NIM was serving, so maybe some of that's on the serving, not the model. but the biggest model NVIDIA has ever made, on the exact task it was designed for, couldn't tell it had built nothing in 100 minutes. same task on a local model, below thread👇.

Sudo su

32,589 görüntüleme • 2 ay önce

According to Roy Scheider, "Filming 'Sorcerer' (1977) made 'Jaws' (1975) look like a picnic". He considered the Bridge scene to be the most dangerous he has ever acted in. The scene was so real that he said, "What the audience will see on that screen is what really happened." Building the bridge cost about a million dollars. It took 3 months to complete. Friedkin called it "A mad enterprise and definitely life threatening". The bridge was built on what they considered to be a "perfect river" in Dominican Republic. It was 12 feet deep. However, there were weeks without rainfall and the level of the river started dropping. The river has never went dry before according to the locals. When the bridge was completed, the river had only 1 feet of water making it impossible to film the scene there. With only 2 scenes left to film, the crew had to leave for Mexico to shoot the bridge scene. So they decided to shoot the scene in a river in an Aztec village in a remote location in Mexico. The Bridge had to be reassembled and anchored in the new location. The shutdown of production had lasted a month which was a huge expense to the management. When the crew arrived at the village, there was a huge exodus of the local population. When Friedkin enquired, he was informed by one of the authorities that it was because of word of William Friedkin's arrival. They were a deeply religious people, and when they heard the man who made "The Exorcist" (1973) was coming to their village: they felt it was "bad karma". But few of the locals stayed and helped with the filming. Before arriving, Friedkin was informed that it rained often in the area, but didn't know that it was only in the summer season. Since it was the fall, there was no rain and the river began to drop at the rate of six inches or more a day. The water was diverted to the location of the bridge in the river using large pipes and pumping equipment. Since Friedkin wanted to shoot the scene in the rain, they brought in half a dozen large sprinklers that drew water from upriver. The scene runs twelve minutes, roughly 10 percent of the final cut, but it took months to complete and cost more than $3 million, most of it not budgeted. ("The Friedkin Connection: A Memoir", William Friedkin, 2013 & Roy Scheider's interview to the The New York Times, 1977) P.S: On this day, 49 years ago, "Sorcerer" (1977) was released in the USA.

DepressedBergman

76,740 görüntüleme • 2 ay önce

In a recent interview with Press Box PR, Raphaël Colantonio, the founder of Arcane Studios, shared his experience of working with Valve on the cancelled Half-Life 2 expansion called "Return to Ravenholm": “It took us a decade to move on from the pain of the Ravenholm project because we were incredibly excited about doing that game and we put in a lot of creative energy into it and eventually when the decision was made to cut it, we did not understand it as it happened. We were young and we were just refusing to see the reality of the situation and the reality was much simpler than the way it looked to us. It all came down to the economics behind Half Life 2 and how the DLCs were selling. They had a business ratio between how much they could spend and how they could make it to be worth it to them because it takes the attention of the entire structure on their side. They had tried to create something internally and then they had tried with an external team already in America but it did not work out because they were too expensive so they worked with us because they liked our work and we worked with them during that time. They thought, you know, why not try these guys? We seemed pretty effective and eventually, like the two other attempts that they had looked at, we went somewhere with it, and it was cool. What we had was really, really cool. But it still needed more time, and too much time to finish. Beyond it being cool there was a business call at some point where it was decided that it didn’t really make sense. It was very painful for us but it made sense, we grew up and got a little bit more mature about it.”

‎Gabe Follower

158,051 görüntüleme • 2 gün önce

anthropic's head of product just revealed how they're able to ship faster than any other AI company. their secret: "side quest maxxing." here's how it works: instead of long-term roadmaps, anthropic runs on unplanned afternoon experiments. anyone on the team gets full freedom to spend an afternoon prototyping an idea and show it to the team. you get to skip the approval process entirely. then, employees at anthropic try it. if they keep using it the next day and the day after that, it gets polished into a real feature. if nobody touches it again, it dies. that's the whole process. claude code on desktop started as one engineer's afternoon project. he wanted it to work on desktop so he built a prototype. people on the team started using it immediately. so they shipped it. the todo list feature started the same way. someone built it, the team adopted it internally, and it became one of the most-used parts of the product. plugins started when one engineer shared a spec with claude code and the prototype that came back was close to production-ready. went from idea to working feature in a single session. they also killed standup meetings. instead of telling people what you're working on, you just show a working demo. all walk no talk basically the team structure makes this possible. > designers ship code. > engineers make product decisions. > product managers build prototypes. everyone can take an idea from concept to working demo without waiting on anyone else. the biggest features at a $380b company came from afternoon experiments that nobody asked for. honestly this matches my own experience cooking with ai. some of the best workflows i use every day came from just fucking around. opening a session with zero intention and asking claude what it can do, or jamming on a random idea to see where it goes. if you're only using ai for tasks you already have in mind, you're missing the best part. open a session with no agenda. ask it to surprise you. try building something stupid. half the time it goes nowhere. the other half it becomes the thing you use most. you need to be sidequestmaxxing.

Ole Lehmann

106,072 görüntüleme • 4 ay önce

This is psychotic behavior. It’s not just malignant narcissism and dementia. He has lost contact with the objective world and expects everyone to live in his alternate reality with him. It’s all manufactured by his own psyche and no amount of evidence will change it. This is a terrified inner psyche constructing a protective world. He’s insane and existentially dangerous. “Jeanine Pirro made a mistake. It was vandalism. I just told you we did, I think, 78. One of the things was to reflecting pool, hasn’t worked since 1922, because it always leaked. It always leaked from 1922. You know, it's the largest longest pool ever, all that. The concept is beautiful, but it always leaked because of the size, because of maintenance, whatever. But it was built in 1922, and from the day it was built, it like, Biden spent $58 million. Barack Hussein Obama spent much more than that. He said, "I have an idea. Let's take the river from, let's take the water from the Potomac." So they took the water from the Potomac, and it was putrid. It was putrid. It was a disaster. Biden was a disaster. I said, "I'm going to get that face." Along with Doug Burgum, Department of Interior, and we worked harder than that. Now we worked harder in all '70s. It's actually, I think, '81 now. You saw the new horses that just opened by the bridge with the gold. That was the way they were many years ago. Now they looked better than they did many years ago. But these monuments looked better than they did when they were originally built. Not just a little fix up. We made them the way they were plus, and everyone's so proud. The biggest, one of the bigger jobs was not really the biggest, believe it or not, but one of the biggest was the reflecting pool. And we did a great job. We sandblasted the stonehouse at his granite, so it has a long life. We did a great job. We got very expensive material to put on top of the surface that always leaked, because it was stone. It was a stone, sir. Always leaked. And we put it on, and it was beautiful. Now we have photographs of tapes, like moving cameras, right? We have them where people are on the side, cutting it with a box knife. So now I'm not saying I was 100% thrilled with the contractor, but the contractor was rushing. We wanted to get it open for July 4th. And we got it done. But in addition, there was mandalism. Number one, look at the grass, where the grass was all knocked out with a very powerful ingredient. We know very well what it is. But like I say, the name, because people get ideas. But they put a terrible phrase, I won't say that either, but a terrible phrase, on a massive piece of grass that we had just replaced. Kill the grass. We had to replace the grass. A lot of grass. In addition, they took knives or cutters, and they cut the material. And they put their hands and they pulled it. And there was also some maintenance things that we would have routinely fixed. But we had that up, and it was perfect. And then you had the bubblers, which get rid of what forms in the water. And that works great, but they turned it up because of the fight. Because it competed a little bit of noise during the fight. They turned it off, and it grows very rapidly in the water. But that actually works very well. And that's all. So now we're just about set to reopen it. And we had to replace cut areas where they cut it and ripped it. But there's a tape. And Jesse Wooders, by the way, did a whole big thing on it. There's a tape that's out there that I posted on truth. Where you have people leaning over the side, the weakest part, because of the flexibility. You know it moves. It's very complicated. It moves. And the softest part has to be flexible. And that's the stuff you can cut. And these guys knew what they were doing. And they cut all the way along the base, cut, cut, cut. Now, when you need to look at it, it was a day after the 4th of July, with the largest fireworks display in the world.”

Jim Stewartson, Decelerationist 🇨🇦🇺🇦🇺🇸

2,263,746 görüntüleme • 27 gün önce

i watched gemma 4 12b build something genuinely impressive today, and then loop itself to death right in front of me. the full run is in the video, sped up but completely uncut, watch it to the end and you will catch the exact moment it stops building and starts looping right in the middle of the work. the task was clean, build a single file gravity simulator, n-body physics, orbits, collisions, running locally on one 3090 through an agent. and for ten minutes it was a joy to watch. it reached for a symplectic integrator on its own, the correct one, the kind that keeps orbits stable instead of spiralling out. real gravity with softening, proper orbital velocities, momentum conserved on collision. the physics was right. the thing actually worked. then on the very last step, writing a few tests to prove its own code, it fell into a loop. not a crash, a loop. it started repeating itself and would not stop. ten more minutes, thirty four thousand tokens into a single answer, the same fragments over and over, until i killed it myself. so it's not that gemma can't code. it did the hard part beautifully. it cannot finish. it cannot hold a long task together without unravelling, and finishing is the entire job in agentic work. here's the part that stings. i run this exact task, same harness, same card, on the chinese open models, qwen especially, and i never see this. they build it, they test it, they stop. every single time. google has the raw capability, you can see it sitting right there in the code, and then the model loops itself to death on a task a 27b from alibaba finishes clean. open weights, apache 2.0, so much to love on paper. i just need it to know when to stop talking.

Sudo su

39,719 görüntüleme • 2 ay önce