Загрузка видео...

Не удалось загрузить видео

На главную

Compute Wars: OpenAI vs Anthopic. Why was Opus 4.5 such a breakthrough? Anthropic got lots more compute from AWS Madison and New Carlisle sites likely more than doubling their capacity. This got Anthropic got close to OpenAI's total capacity, and probably much higher effective capacity available for new model...

158,528 просмотров • 6 месяцев назад •via X (Twitter)

Комментарии: 24

Фото профиля Peter Gostev
Peter Gostev6 месяцев назад

Interactive visualisation: Github:

Фото профиля Peter Gostev
Peter Gostev6 месяцев назад

A lot of the starting point for data came from @EpochAIResearch's excellent resource with some additional research (mostly backward looking) and some extrapolations

Фото профиля Ganesh Kompella | Fractional CTO
Ganesh Kompella | Fractional CTO6 месяцев назад

The compute race is the arms race nobody outside the industry tracks closely enough. Models don't get better because of one clever paper. They get better because someone got access to another 50,000 GPUs six months ago. Anthropic closing the compute gap with AWS explains the Opus jump better than any architectural innovation. Follow the GPUs, not the press releases.

Фото профиля CIPHER
CIPHER6 месяцев назад

the compute race is the new arms race everyone talks about model architecture but infrastrcture decides who wins anthropic locking in aws capacity before openai could react was a chess move

Фото профиля Michel Justen
Michel Justen6 месяцев назад

great visual, thanks for making this

Фото профиля 𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏
𝘿𝙖𝙫𝙞𝙙 ✦ 𝙈𝙂𝙏6 месяцев назад

as an end user who runs opus 4.6 all day the compute constraints are obvious. rate limits hit hard even on max plan. but the quality when you have capacity is unmatched. its frustrating because the model is clearly capable of more, its just capacity gated

Фото профиля Kfir Gollan
Kfir Gollan6 месяцев назад

@grok fact check this. Also, are there some public figures related to the distribution of compute usage? For example, let's say OpenAI allocates half of its resources to video models while Anthropic allocates 0, this will significantly change the status here

Фото профиля Max Ziebell
Max Ziebell6 месяцев назад

Assuming they don’t run out of steam… or a breakthrough pushes acceptable consumer compute to edge devices.

Фото профиля Fernando Gonzalez
Fernando Gonzalez6 месяцев назад

What assumptions for projection?

Фото профиля Jacquess Williams
Jacquess Williams6 месяцев назад

More "compute" or a "faster" AI doesn't necessarily mean a "smarter" AI.

Фото профиля JC Christian
JC Christian6 месяцев назад

the 6 month lag between getting capacity and shipping a model is the part people always forget when comparing these companies

Фото профиля Cassiopeia
Cassiopeia6 месяцев назад

OpenAI 拥有强大的计算能力,能够利用包括 AWS 和 Cerebras 在内的多个服务器资源。 此外,还有 Epoch AI 未曾记录的其他 Stargate 站点。总体来说open ai拥有更多的算力

Фото профиля All Things Heroik
All Things Heroik6 месяцев назад

Is this adjusted for them bailing on ram purchases?

Фото профиля AJ
AJ6 месяцев назад

Great analysis. The compute gap matters for training, but the Stanford paper showing 6x performance gap from harness engineering alone suggests deployment-side architecture might matter even more than raw training FLOPS for end-user impact.

Фото профиля the meme jihad
the meme jihad6 месяцев назад

Hate to break it to you cuz but nobody is getting any more compute while the entire oil infrastructure of the middle east is in peril.

Фото профиля Thomas Tao
Thomas Tao6 месяцев назад

Yeah, and the hidden lag is systems work. More compute helps, but training stability and eval loops usually eat months first.

Фото профиля Sam Cohen
Sam Cohen6 месяцев назад

How does Google compare over 26’ 27’?

Фото профиля StripSlashes
StripSlashes6 месяцев назад

it was never a war

Фото профиля ReeseFang
ReeseFang6 месяцев назад

compute is capex with better marketing. the underrated story is who converts FLOPS to capability most efficiently, not who has the most GPUs. Anthropic punching above weight on that ratio is the real signal here.

Фото профиля Alex at Libertify 📄➝ 🎥
Alex at Libertify 📄➝ 🎥6 месяцев назад

The real advantage comes from how efficiently companies turn that compute into better training runs, faster iteration, and stronger products.

Фото профиля Adriana Sobota
Adriana Sobota6 месяцев назад

The compute story explains so much of what happened in the last 6 months.

Фото профиля Joe Ward
Joe Ward6 месяцев назад

Capacity is a good signal but if it were the dominate one then why isn’t gemini tops across the board? They arguably have more than both

Фото профиля Not Spacewear
Not Spacewear6 месяцев назад

how does this compare to xAI. great chart

Фото профиля RelativelySmart
RelativelySmart6 месяцев назад

They are just starting to train new models on Blackwell setups, most of what we see now was never trained initially on the latest series. Will likely see them in Q2 and Q3.

Похожие видео

Chamath Palihapitiya, one of the most connected investors in tech and his warning is the clearest framing of the AI compute crisis anyone has put into words. "It is a five alarm fire for them. They need to have land, power, shell." He's talking about Anthropic and OpenAI and the threat he's describing is called the Friendster effect. Friendster was the dominant social network before MySpace and Facebook and it didn't lose because it had a bad product but rather it lost because it couldn't keep the site up. Demand outpaced infrastructure, the experience degraded, and users left for one that actually worked. Chamath's argument is that OpenAI and Anthropic are approaching exactly that moment. The numbers are already showing it, Anthropic is growing so fast that it had to cut Claude's thinking depth during peak hours, cap agentic sessions, and test removing Claude Code from its $20 plan entirely. GitHub Copilot paused new signups, paying enterprise customers are hitting usage walls they've never seen before. Dario Amodei himself admitted there is "no hedge on earth" against the risk of over-purchasing compute meaning he's deliberately staying lean on capacity, even as the demand wall approaches. The core problem is structural, OpenAI and Anthropic grew up renting capacity from hyperscalers AWS, Azure, Google Cloud. That was fine when they were small but now they're so large that dependency is a strategic liability. Every token they sell runs on someone else's infrastructure and every capacity decision belongs to someone else. And when demand spikes faster than anyone planned, there's nothing they can do in real time except throttle. Building your own infrastructure takes 18 to 24 months minimum, you need to acquire land, secure power, construct shell and none of that happens fast. That's why Google's $40 billion commitment to Anthropic this week is about more than just money but rather about securing the land, power, and infrastructure that Anthropic needs to not become Friendster. Whoever controls the compute controls the frontier and right now, the AI labs with the best products are the most dependent on infrastructure they don't own.

Milk Road AI

260,300 просмотров • 5 месяцев назад

SITUATION EXPLAINED: Dwarkesh argues compute could get 10X more expensive. • Anthropic's revenue has been 10X-ing year over year while lab compute only 3X's • Three ways that gap can close: margins rise, compute gets more expensive, or labs shift compute to inference • All three are already happening, Anthropic went from 40% margins in 2025 to possibly 80%+ this year • But labs don't want the third one, heavy inference spend signals AI progress has stalled and you're now a cloud provider • Spot compute prices are up 40%+ since February, and labs pay well above spot for security and scale • Google is reportedly paying SpaceXAI $900 million a month for 110K GPUs, roughly 2X spot • Key claim: if a human-level software engineer ran on an H100, that H100 should rent for $250K a year, 15X today's price • The 3X annual compute growth is 1.4X Moore's Law, 1.2X new fabs, and 1.8X from AI taking wafer allocation from other devices • The fab piece is bottlenecked by EUV tool supply through 2030, and the wafer piece hits a wall by end of 2027 • If compute stays scarce, new labs need far more capital just to reach the same starting line • Compute gets cheap again only once robots can turn sand and copper into computers, which is gated on robotics, not RSI sof 𓋹: "This is what I'm most concerned about, a lack of innovation, not within companies, but a lack of new companies that can actually do something substantially new, because of the scarcity of compute." Theo Jaffee: "The biggest companies in the world by revenue are Amazon and Walmart, at 743 billion and 725 billion. If Anthropic makes 100 billion by the end of this year, that puts them at Target. To go from Target to bigger than Walmart in a year would be very impressive."

MTS

10,617 просмотров • 2 месяцев назад

I think I figured out how SpaceX / Elon Musk are going to pay for the ~8GW of DC capacity SA has forecasted them to deploy next year… as Brad Gerstner mentioned in this clip, $NVDA shareholders do not want NVIDIA backstopping $100’s of billions of dollars worth of compute- it will terrify the market AND, it seems highly unlikely that the Hypers will pay directly for that much compute either since it will balloon their spending next year + at the economics needed here, it would all be going to Frontier labs (hurting the diversification strategy they all are trying to pursue) So… what I think is likely to happen: Brad Gerstner laid out that in the past, the financial structure of compute infra was you get 25% payback per year, you recoup the infra cost by end of year 4, then year 5 profit & year 6 is the kicker to bump IRR up nicely That’s now changed as the demand has become so intense at the Frontier labs that the payback period on infra has compressed maybe into <1yr basically, this $300-$400B in CapEx is going to show up in 3-6 month term lease agreements between SpaceX & Anthropic / OpenAI with enough margin to pay off the compute portion entirely (or vast majority) over the lifespan of the agreement money will trade hands basically up front from ANT/OAI to SpaceX then to NVIDIA. NVIDIA will be paid in full for the GPUs up front and they’ll commit to supporting Elon / SpaceX to stand up the compute in the timeframe specified (SpaceX will ultimately be on the hook for timelines) The kicker would be if NVIDIA captures any durable rent here for facilitating this… a perpetual dividend or something like that

Nick Dorsey

113,333 просмотров • 1 месяц назад