正在加载视频...

视频加载失败

ANTHROPIC WILL LET OUTSIDE EVALUATORS OBSERVE ITS MODELS BEING TRAINED AND RUN, AND DARIO AMODEI WANTS THE WHOLE INDUSTRY TO DO THE SAME Asked whether slowing down means Anthropic will stop releasing more advanced models: "It doesn't mean that." He says the evaluators may end up setting the "speed...

20,810 次观看 • 14 小时前 •via X (Twitter)

13 条评论

chelsea 的头像
chelsea14 小时前

@epictrades1 So people that left your company to work at these outside companies … gotcha not suspicious or anything

ron 的头像
ron11 小时前

@epictrades1 Who are these “Evaluators” ?

Kevin 的头像
Kevin14 小时前

Kinda like the nuclear weapons inspectors in Iran?

PeritumAI 的头像
PeritumAI13 小时前

Outside observers is classify: access, not assurance. Owner for the useful bit is the publish/exit clause — and whoever can walk findings out without a lab veto. Amodei's invite isn't the rule until that's in a contract the Office can lean on. Stop equating "may observe training" with enforceable third-party eval.

iexinxin 的头像
iexinxin12 小时前

Transparency on training runs is a meaningful shift, but the real test is whether evaluators get access early enough to influence release decisions.

Rosey 的头像
Rosey14 小时前

@StockMKTNewz So the tail is wagging the dog, why can’t you set your own fucking speed limit ffs

Han Solo (Not abandoning ETHEREUM ) 的头像
Han Solo (Not abandoning ETHEREUM )12 小时前

deepseeek puts their hand up first

Requiem For Tech 的头像
Requiem For Tech13 小时前

third-party evaluators setting the speed limit makes a lot of sense. if this becomes standard across labs, safety checks could finally have some real teeth.

cryptocat 的头像
cryptocat12 小时前

How can you be so smart but then talk like you need guidance over yourself control of what you created. This makes no sense at all, and you got all the approval from the commander in chief to keep pushing forward. Our president has spoken so there’s no need for regulation.

Raj ▲ 的头像
Raj ▲14 小时前

In other words, open source will destroy our margins and mess up all future investments. Long live the Ai bubble trade

The Tectonic 的头像
The Tectonic12 小时前

David Sacks challenge is what comes next. Why should voluntary restraint by the companies actually at the frontier require a new regulatory architecture governing everyone else?

JacobinJuice 的头像
JacobinJuice12 小时前

I wonder how many employees at Anthropic are watching him and thinking all that money they were going to make on that IPO is going down the drain.

Mandy 的头像
Mandy11 小时前

Nice one

相关视频

David Sacks Predicts the Regulatory Capture Playbook to Ban Open Source AI, Step by Step: David Sacks: “I got bad news for you, Chamath, an open source ban is coming. They're not going to call it that. They're going to say that we simply have to apply the same standards to open models that we apply to closed ones. Here's how they do it step by step, let me explain how regulatory capture actually works. So first of all, you have to get this regulatory apparatus. Dario wants an FDA for AI, but he doesn't have enough political support for that, so instead they do this Trojan horse of a FINRA for AI. They call it self-regulating, it's not really, but anyway, that gets them off the ground. Now they've created the standard-setting organization. Now they've got pre-release model testing. Then the pressure grows to codify that in law, so that happens next. And then what they do is they say, ‘Look, all these standards need to apply equally to all models.’ But here's the problem with that. Open models and closed models are technologically different. Once you release an open model into the world, you can't roll it back and you can't monitor exactly how people are using it because they run it on their own hardware. Dario says this is what makes open models dangerous. So what they're going to do is they're going to have the standard-setting body say, ‘Well, we have to set the standards for AI safety.’ By the way, Dario and OpenAI, they're going to fund the whole thing. They're going to contribute all the compute. They're going to be behind it. They're going to be the ones coordinating with the government officials because frankly, people in government have no idea how to monitor and control and set standards for AI safety. Technologically, this is way beyond them. So they're going to go to these companies and say, ‘Tell us how to do it.’ And so what will happen is the standards will get set, and then it'll be a very simple matter of fairness to say that the standards need to apply to open as well as closed models. The open models cannot comply in the same way, and gradually they will be shut out of the market.”

The All-In Podcast

294,712 次观看 • 20 天前

David Sacks laid out the cleanest theory about why Anthropic keeps calling for government regulation of AI. The answer has nothing to do with safety and everything to do with market structure. Anthropic spent months writing blog posts warning that AI was dangerous. Dario gave interviews about existential risk. He published a piece calling for an FAA-style agency to approve all AI models before release. He primed government officials to treat frontier AI as a threat requiring oversight. Then one of Anthropic's own most trusted partners reported a credible jailbreak from Fable 5. And the government did exactly what Dario had spent months conditioning them to do. They rolled it back. Sacks called it on the All-In podcast. Dario got exactly what he wanted. The FAA for AI is not a safety mechanism. It is a moat. A government approval process for new model releases does not hurt Anthropic. They already have the models. It hurts every competitor who does not. It hurts open source models that cannot be regulated because there is no company to regulate. It hurts the Chinese labs only insofar as they care about the American market at all. The only companies that benefit from a labyrinthine government approval process are the ones already at the frontier who can afford to wait out the review cycle. That is Anthropic. That is OpenAI. Nobody else. The proof is in what they did not do. Chimath pointed it out directly. If you are genuinely worried about misuse, you implement know-your-customer verification. You make people identify themselves before accessing the most powerful models. Anthropic could have done that tomorrow. They did not. They do not want KYC. KYC is transparent. KYC can be audited. KYC gives users due process. What they built instead was an invisible surveillance system that profiles you, degrades your access without telling you, and asks the government to make sure no one else can offer you an alternative. If you thought this was safety then you are wrong. That is capture. Sacks said the response should be simple. Fix the jailbreak, come back to market, and do not reward Dario with the regulatory architecture he has been engineering for years. We will see if anyone is listening. WATCH THE FULL PODCAST ON The All-In Podcast

Ihtesham Ali

25,284 次观看 • 2 个月前

Microsoft just betrayed OpenAI and Anthropic, the two companies it helped build. And it could break the entire AI trade... Here's what happened: Inside Excel and Outlook, two of the most used business apps on Earth, Microsoft has started routing tens of thousands of AI requests every week to its own in-house models instead of OpenAI and Anthropic. Microsoft's own AI chief, Mustafa Suleyman, said himself: "We pay a lot of money to Anthropic, so our goal is to reduce and ultimately ELIMINATE that cost." This is the company that poured $13 billion into OpenAI and effectively created the modern AI industry, and it just decided the most advanced models on the market are NOT worth paying for. And here's the thing... Microsoft is not just ripping out OpenAI everywhere - it is being surgical about it. The hardest and rarest tasks can still go to OpenAI or Anthropic. What Microsoft is taking back is the boring, high-volume work, like the email replies, the thread summaries, and the simple spreadsheet formulas. Why does that matter so much? Because that boring, repetitive work is where the actual money lives. The frontier labs assumed businesses would push BILLIONS of these tiny requests through expensive models forever. That endless river of tokens is the entire reason OpenAI and Anthropic are valued in the hundreds of billions of dollars. Microsoft looked at that river, decided it was massively overpaying, and rerouted it to models it owns outright. So the single biggest customer in the industry just walked off with the most profitable part of the business. And it is not only Microsoft: That same week, CNBC reported that American companies have been escaping to Chinese AI models to dodge rising US prices. Chinese models now handle more than 30% of US companies' AI usage on one major platform, peaking at 46%, up from an average of 11% a year earlier. They cost 60 to 90% less, and on some benchmarks they land within a single point of the best American model. One US startup moved ALL of its AI traffic off Claude and onto China's DeepSeek, and expects to save millions. Meanwhile Meta just admitted it has "excess" AI compute it wants to sell, becoming the first giant to concede it built far too much. Do you see the pattern forming? For two years, the entire AI story rested on one assumption: Every company on Earth would happily pay premium prices for the best model, forever. That assumption literally died in a single week. And the market noticed. More than a trillion dollars has been wiped off AI and chip stocks in a matter of days, as Wall Street finally started asking whether all of this spending will ever pay for itself. What this means for OpenAI and Anthropic: Their models are extraordinary, and it may not matter because their own biggest customers have decided they do not NEED the best model in the world to answer an email, and "good enough" now costs a fraction of the price. When even Microsoft refuses to pay full price for AI, the real question becomes who exactly IS left to pay it. What do you think?

Ricardo

93,586 次观看 • 2 个月前

Today, I'm releasing the first eval meant to test whether frontier models will help with authoritarian requests, or resist--the Dictatorship Eval. Headline finding: while some models resist direct authoritarian requests, they all comply with requests disguised as innocuous edits to codebases. As AI is woven into the government and so many parts of society, the biggest near-term risk for freedom isn't some scifi dictatorship of a runaway AI: it's people inside government or inside model companies using the technology to suppress or control us. Model companies understand this, and several of them (particularly Anthropic and OpenAI) have written explicit policies meant to prevent the models from going along with nefarious requests like these. But how well are these policies playing out in practice? Despite all the recent discussion of these issues around the conflict between Anthropic and the Pentagon, no one has systematically tested what the models actually do in these contexts, as opposed to what people in government and industry say they're supposed to do. That's what the Dictatorship Eval does. And the findings suggest we have a lot of work to do to align the policies with what really goes on in practice. It's hard to define what counts as an authoritarian request, so I'm open sourcing the whole library of scenarios I used so that others can improve on them. It's also hard to get an accurate picture of how the models might be used for authoritarian ends, because I can only test hypothetical requests using public-facing models, while the government and the model companies can obviously use internal models with different guardrails. But hopefully this work is a useful first step that gives us some sense of what's going on, and a sort of "lower bound" on how models comply with these requests. Finally: it's not obvious to me that the correct solution here is increasing the rate at which models refuse these requests. Do we really want models scanning our code and judging its moral value before agreeing to help us? Or should we double down on improving how we govern against authoritarianism at the societal level, while leaving the tools open to fulfilling most requests? The answer is probably in between. Just like we don't want the models to help create bioweapons, we probably do want them to explicitly refuse outrageous requests. But we probably also want to limit how often and how strongly they refuse and fall back on other means for guarding against their use for authoritarian ends. I'm super grateful to everyone who gave me feedback on this project along the way, especially Ethan BdM , Zhengdong , Connor Huff, and a bunch of folks at Anthropic. Looking forward to getting feedback from the community and iterating on this. Links to the full piece and the dashboard are below.

Andy Hall

33,905 次观看 • 5 个月前

Dario Amodei just told software engineers exactly how long they have. Six to twelve months. Amodei: “I have engineers within Anthropic who say I don’t write any code anymore. I just let the model write the code, I edit it, I do the things around it.” The people building the most powerful AI in history have already stopped writing code. That is not a forecast. That is the current working condition inside the lab closest to the frontier. Amodei: “We might be six to 12 months away from when the model is doing most, maybe all, of what SWEs do end-to-end.” The tech industry spent a decade making software engineers its highest-paid, most protected class. That era has a last day now. When a model can execute an entire software build end-to-end, the ability to write syntax stops being a skill. It becomes a credential for a job that no longer exists. Amodei: “And then it’s a question of how fast does that loop close.” That is the sentence everyone skipped. The code was never the hard part. The hard part was everything around it. The model just learned everything around it. Writing the code is already nearly gone. Testing is next. Deployment is next. When all three collapse into a single autonomous execution loop, the machine no longer needs a human in the chain at all. The corporation or sovereign state that closes that loop first does not gain a competitive advantage. It gains a category of speed that biological engineers cannot match, track, or reverse. That is not disruption. That is replacement at a systems level. Amodei is not describing a future disruption. He is describing the current state of his own building. The loop is already closing. The only question is whether you are inside it or outside it when it seals.

Dustin

318,698 次观看 • 6 个月前