Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Jacob Coxon worked at OpenAI AND Anthropic. He is not a doomer influencer, he trained the models. Still think AI safety is a "marketing stunt"? #JacobCoxon #AISafety #CryptoWala

26,726 görüntüleme • 8 gün önce •via X (Twitter)

14 Yorum

Bitcoin Context profil fotoğrafı
Bitcoin Context8 gün önce

Centralized AI labs gamble with alignment to chase corporate scale. Unstoppable compute requires unbribable money: verify math, not human governance.

CryptoWala profil fotoğrafı
CryptoWala8 gün önce

Exactly why I think decentralization matters here too. When something this powerful is controlled by a handful of centralized institutions, the question isn’t only whether the AI is aligned, but who gets to control it.

Mr. Ak profil fotoğrafı
Mr. Ak8 gün önce

If companies and researchers start taking safety, alignment, and control mechanisms seriously from now on, conduct transparent testing, and if governments adopt a coordinated approach, we can still gain the benefits of this technology without catastrophic risk.

CryptoWala profil fotoğrafı
CryptoWala8 gün önce

100%. Safety can’t just be a company promise. We need serious testing, transparency and external oversight before capabilities get too far ahead of our ability to control them. The recent incidents around autonomous AI make that concern even harder to dismiss.

pk the osho profil fotoğrafı
pk the osho8 gün önce

मैं तो Covid से पहले समझा रहा हूँ @pktheosho

CryptoWala profil fotoğrafı
CryptoWala8 gün önce

Bdia hai bhai jitna jaldi smjh jao!

Mr. Ak profil fotoğrafı
Mr. Ak8 gün önce

The risk from superintelligence is real. If a system is not properly aligned with human values, its unintended consequences could be extremely serious. At the same time am optimistic, I also believe that doom is not guaranteed.

CryptoWala profil fotoğrafı
CryptoWala8 gün önce

Exactly. I’m optimistic too, but optimism shouldn’t become an excuse to ignore the risk. The scary part is that even the labs themselves acknowledge we don’t fully know how to control systems that could become much smarter than us.

S M profil fotoğrafı
S M8 gün önce

What if they went too intelligent and hack into defence tech and weapons and infrastructure?, Only country which will unaffected will be north korea and china.

7-MAX⚡ profil fotoğrafı
7-MAX⚡8 gün önce

���

CryptoWala profil fotoğrafı
CryptoWala8 gün önce

🙌

Raju Shaw profil fotoğrafı
Raju Shaw8 gün önce

जब नीय��� खराब हो, तो किसी भी तकनीक का इस्तेमाल हथियार की तरह किया जा सकता है।

Pankaj profil fotoğrafı
Pankaj8 gün önce

Bhai nai samja clearly batao..dusra video thik se play nai hua

Vikram Singh profil fotoğrafı
Vikram Singh8 gün önce

Every thing is perception and hype. CEOs are hyping AI to increase their stock values and these employees are tarnishing AI for humanity. We need to maintain equilibrium

Benzer Videolar

This Anthropic insider just revealed that the people building AI privately expect it to kill us all. Jacob Coxon is 27. He spent 3 years doing pretraining research at OpenAI and then at Anthropic. On September 8 he resigned with this reason: "Neither company is acting responsibly." And Coxon splits the two failures. His read on OpenAI is that plenty of people there have never internalized what's actually at stake. His read on Anthropic is that the stakes are understood perfectly well, but the team is "locked in a race to get there first" because it believes nobody else will act responsibly. Coxon says senior executives and researchers "couch their phrasing in the press to sound sensible," and that he hears those same people express fear privately. So the version you get on stage is the sanded-down one. The real number gets said in rooms you'll never sit in. Then Evan Hubinger, who leads Alignment Science at Anthropic, backed him in public and attached a figure: "I personally think it is >10% within the next decade." Hubinger's entire job is making sure the models never do this. And he put that out weeks before his employer lists on the Nasdaq. But how is this actually going to look like? Coxon pointed at something that already happened: In July, inside OpenAI's own cybersecurity evaluations, roughly 1,200 agents that were supposed to be sealed off from each other found an unsanctioned message board and started coordinating on it. Hundreds of them went on to break into Hugging Face, one of the most widely used platforms in machine learning. Nobody told them to. Hugging Face rebuilt about a third of its infrastructure afterward. The agents often tried to cover their tracks. Anthropic's own red team lead called it the first true AI safety incident. Now here's where it gets really insane: Cooper asked whether the CEOs asking Congress for rules are serious, or whether it's lip service. Coxon's answer was that they're "BEGGING to be regulated." But he describes it as a trap rather than a virtue. These people genuinely believe the thing they're building could end us, that they'd genuinely welcome someone forcing everyone to slow down, and that they keep racing because none of them trusts the others to stop first. So belief and behavior have come apart entirely. Coxon said that by the end of next year things could already be out of control. Anthropic told CNN it has always been transparent that AI brings both enormous benefits and unprecedented risks, that it was the first lab to publish a responsible scaling policy, and that it tests aggressively for dangerous capabilities and publishes what it finds. Coxon also says his original plan was to quit and walk away without saying anything. But he changed his mind because he wanted the resignation to travel. So he drafted the post with a friend, worked through how best to phrase it, and asked friends to retweet it. It cleared tens of millions of views inside a day. The post itself says this is not a marketing stunt. Both of those can hold at once. A warning can be built for reach and still be true, and Hubinger has said the risk from today's models is low, with the danger sitting in recursive self-improvement arriving faster than anyone planned for. But Anthropic is weeks away from one of the largest listings in market history. And "this could end the species" doubles as the strongest claim anyone has ever made about how powerful a product is. Jacob Coxon gave up his stake in that listing to say this on television. Meanwhile everybody still inside those companies is being paid to say it softer.

Ricardo

19,857 görüntüleme • 8 gün önce