Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Open-source models are already taking the easier tasks. The bull case for frontier labs rests on one question. "A lot of software engineering, white-collar work in general, does not need Fable 5.1 or Astra 6 level intelligence." "A lot of businesses, especially lower gross margin businesses, are very rationally...

35,742 görüntüleme • 2 gün önce •via X (Twitter)

18 Yorum

SemiAnalysis profil fotoğrafı
SemiAnalysis2 gün önce

Get the Full Podcast on Spotify:

Anderson profil fotoğrafı
Anderson2 gün önce

Frontier labs selling genius while most work needs a cheap model

Deep profil fotoğrafı
Deep2 gün önce

which tasks do you count as easy for the open models?

Stevan Boljevic profil fotoğrafı
Stevan Boljevic2 gün önce

When you strip out API margin differentials (source - your numbers) the frontier labs have lower compute cost per task for easier tasks. With additional compute supply they could capture the market for these tasks too with lower margins on Luna/Sonnet/Terra class models There’s plenty of upside being lowest cost producer of commodity factor of production - look at the Gulf

Macro Bombastic profil fotoğrafı
Macro Bombastic2 gün önce

tbh the easy task money was never the frontier labs game anyway

ALEX SERRA profil fotoğrafı
ALEX SERRA2 gün önce

Besides cybersecurity,what other AI niches could open models attract in the future ?

l-tzhar-bijaz profil fotoğrafı
l-tzhar-bijaz2 gün önce

thank you very much. how about frontier labs are going to dominate every single point of the pareto cap / price curve?

Joshua Week profil fotoğrafı
Joshua Week2 gün önce

And a lot of compute doesn’t need a power hungry GPU. Please follow these examples

Alex Crișan profil fotoğrafı
Alex Crișan2 gün önce

Well said; but electricity will skyrocket

Eon Vale profil fotoğrafı
Eon Vale2 gün önce

what counts as an easy task keeps moving up every six months, so frontier labs are basically selling a head start that keeps getting shorter

Thomas DiFazio profil fotoğrafı
Thomas DiFazio2 gün önce

Two-thirds of large businesses now pay for AI per Ramp data but the next leg of adoption may be depth: learning the difference of when to use a top-tier model vs. a less intensive one.

RowdyGoose profil fotoğrafı
RowdyGoose2 gün önce

Semianalysis has 0 credibility any more

StockTake profil fotoğrafı
StockTake2 gün önce

Feels like lazy analysis. Open source taking over simpler tasks doesn’t necessarily weaken the frontier labs. It could massively expand the market while pushing frontier models to problems that weren’t economically feasible to solve before. The bigger question is whether the labs can maintain pricing power as intelligence gets cheaper. That’s what will determine who captures the value, not how many hypothetical PhDs the economy can absorb.

Nacho profil fotoğrafı
Nacho2 gün önce

Once an app lets users pick, routine work drifts to the cheaper model, and the frontier labs have to earn their price on the hardest tasks

BullBear.News profil fotoğrafı
BullBear.News2 gün önce

the transition point is when local inference cost drops below the coordination overhead of external APIs

Sterling profil fotoğrafı
Sterling2 gün önce

Unless you're a Big Tech SWE, the cost effectiveness of open source tokens is really really hard to ignore.

Nessuno profil fotoğrafı
Nessuno2 gün önce

See Wallis/North 1986, or our book forthcoming @PalgraveEcon for an unqualified, No.

Tim profil fotoğrafı
Tim2 gün önce

Competition at any level (state/country, corporation, individual) ensures that there is a market for ever higher levels of intelligence. You want to be smarter than your competitor.

Benzer Videolar

Baseten Head of AI Model Training Charlie O'Neill says the future is many specialized LLMs dedicated to specific tasks, with bigger labs deployed on the frontiers of areas like science and math: "People are thinking about intelligence capabilities in the wrong way. People are thinking about intelligence relativistically. They say, 'OK, the open-source gap is like 6 months behind closed-source, and GLM 5.3 is as good as Opus 4.8,' or whatever." "The best way to think about what models can do for you, and for the world, is in an absolute sense." "So for any given task that you want to do with an LLM, there's some intelligence threshold where below that you can't do the task, and above that you have very diminishing returns to more intelligence on the task." "So when you think about it that way, the game of LLMs over the last 5 years has been, 'OK, we have these things we want to do with them. Closed source hits it first... but open-source can eventually do that task. And then for many reasons, once you have the base level of intelligence required to do it, you probably do want to swap to open-source." "It's not really about the [frontier lab] God model being better. Like, if I'm filing a tax return, there is a limit to how much intelligence I need to do that particular thing." "So I think the world is going to look like — frontier closed-source labs are going to continue to push the frontier. You do want to use the most intelligent model. You have very inelastic demand for intelligence when you're doing frontier science or frontier math." "But for a lot of the economically valuable things, it looks a lot like, 'I'm a Cursor, or I'm one of these big companies who are realizing I can't just be a wrapper anymore. I've been through the life cycle of building a product that people love. And I should be using that information to make my model better at the things that I care about, and not at anything else.'"

TBPN

50,531 görüntüleme • 1 ay önce

David Sacks says companies are trapped paying OpenAI & Anthropic because they can't figure out how to use open source models "I think enterprise CTOs would like to shift their token consumption to cheaper models for the obvious reason that it would be more efficient. They are seeing compute costs or token costs skyrocket right now, so everyone's trying to figure this out." "You also have the AI sovereignty issue that Alex Karp talked about. They're worried about giving up the secret sauce or the alpha in their business to a frontier lab that may one day be competing with them. "The problem is, I think in most cases, they don't have the technical ability to do it. Coinbase figured out how to do it. DoorDash figured out how to do it. They built a token routing system that allows them to send frontier tasks to frontier models and non frontier tasks to more mundane models. But I don't think your average enterprise has the technical capability to do that." "This is why the share of wallet of closed models, it actually increased. I think that open source went from 19% last year to 11% this year. So open source as a share of enterprise spending is actually decreasing." "I don't think that means usage is decreasing. I think usage is skyrocketing. It also may be the case that because the whole point of using an open model is you just pay for the compute costs, you don't have to pay a lab, so it may be that it's hard to measure that usage in terms of spend." "But nonetheless, anyone who's saying that these closed models are going to lose or are somehow losing, you're just not seeing it in the data."

dnap

141,614 görüntüleme • 3 ay önce

learned a lot from this conversation with Simon Mo and Matt Bornstein. biggest takeaways for me: -there are a lot of reasons why we should like open-weight models. a lot of these arguments stop at handwavy things like "what if the labs stop releasing frontier models to the public" or "it's lower cost." but simon's position as lead maintainer of vLLM and CEO of Inferact give him authority to talk about some of the other, more interesting and concrete reasons to pay attention to open-weight models, namely that they allow end-users to calibrate latency / other performance metrics with way more customizability than what any of the frontier closed-source labs offer (and without the fear that your job might be met with a refusal at some random point where you're deep in a 2 hour job) -re: the above point...for this reason, a lot of US companies (inferact included!) choose to use open-weight models over their closed-source alternatives. this also isn't limited to internal workloads / research - on a recent a16z podcast the team at Decagon spoke about how something like 90% of their customer service ai agents run on open-weight models that they've fine-tuned. -we should really appreciate how many companies/teams came out researchers fascinated by the wave of very small open-weight models that were being distilled from e.g. gpt-3.5 and earlier models in 2022/2023 (prior to the release of chatGPT!). these small models motivated the development of pagedattention, which then led to vlmm/inferact (at other layers of the stack with similar origin stories, you can look at teams like openrouter or ollama). in other words, we have open-weight models to thank for a bunch of the orchestration infra we now rely on. i think yet another, indirect, way we can point to open-source/weight infra pushing the frontier forward. anyway, a lot more in this convo, it was a lot of fun!

Elena

12,922 görüntüleme • 2 ay önce