Loading video...

Video Failed to Load

Go Home

OpenAI's Noam Brown says that while AI model performance scales roughly equivalently with more training or inference compute, the cost of inference is on the order of 100 billion times cheaper

250,942 views • 2 years ago •via X (Twitter)

9 Comments

Tsarathustra's profile picture
Tsarathustra2 years ago

Source (thanks to @curiousgangsta):

Tom Shafron's profile picture
Tom Shafron2 years ago

that's an odd comparison... one is a marginal cost and the other is a fixed cost. That's like saying a physical store is 500000x more expensive than an item sold in it... yeah but so what, cost of goods is likely a larger expense.

Zacchary Hulsman's profile picture
Zacchary Hulsman2 years ago

Essentially, test time compute distributes the cost of compute, rather than concentrating it in open AI’s hand

Prashant's profile picture
Prashant2 years ago

As the technology progresses , AI models will become smaller , more efficient and better in performance. All this in parallel to cost going down

CM's profile picture
CM2 years ago

but if you serve 1Bn request a day.... what happens post day 100?

Mike Chaves's profile picture
Mike Chaves2 years ago

Fascinating insight from Noam Brown! The scaling of AI performance vs. inference cost shows how much potential there is to optimize efficiency in future models.

Omar Nomad's profile picture
Omar Nomad2 years ago

wow, thanks for sharing, súper interesting talk!

ÐAOrathustra's profile picture
ÐAOrathustra2 years ago

Hopefully inference chains of though gets more and more parallel though.

Ant A's profile picture
Ant A2 years ago

Inference Chips: @GroqInc 👍 @CerebrasSystems 👍 @SambaNovaAI👍

Related Videos