Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Open-source models are already taking the easier tasks. The bull case for frontier labs rests on one question. "A lot of software engineering, white-collar work in general, does not need Fable 5.1 or Astra 6 level intelligence." "A lot of businesses, especially lower gross margin businesses, are very rationally...

35,742 Aufrufe • vor 2 Tagen •via X (Twitter)

18 Kommentare

Profilbild von SemiAnalysis
SemiAnalysisvor 2 Tagen

Get the Full Podcast on Spotify:

Profilbild von Anderson
Andersonvor 2 Tagen

Frontier labs selling genius while most work needs a cheap model

Profilbild von Deep
Deepvor 2 Tagen

which tasks do you count as easy for the open models?

Profilbild von Stevan Boljevic
Stevan Boljevicvor 2 Tagen

When you strip out API margin differentials (source - your numbers) the frontier labs have lower compute cost per task for easier tasks. With additional compute supply they could capture the market for these tasks too with lower margins on Luna/Sonnet/Terra class models There’s plenty of upside being lowest cost producer of commodity factor of production - look at the Gulf

Profilbild von Macro Bombastic
Macro Bombasticvor 2 Tagen

tbh the easy task money was never the frontier labs game anyway

Profilbild von ALEX SERRA
ALEX SERRAvor 2 Tagen

Besides cybersecurity,what other AI niches could open models attract in the future ?

Profilbild von l-tzhar-bijaz
l-tzhar-bijazvor 2 Tagen

thank you very much. how about frontier labs are going to dominate every single point of the pareto cap / price curve?

Profilbild von Joshua Week
Joshua Weekvor 2 Tagen

And a lot of compute doesn’t need a power hungry GPU. Please follow these examples

Profilbild von Alex Crișan
Alex Crișanvor 2 Tagen

Well said; but electricity will skyrocket

Profilbild von Eon Vale
Eon Valevor 2 Tagen

what counts as an easy task keeps moving up every six months, so frontier labs are basically selling a head start that keeps getting shorter

Profilbild von Thomas DiFazio
Thomas DiFaziovor 2 Tagen

Two-thirds of large businesses now pay for AI per Ramp data but the next leg of adoption may be depth: learning the difference of when to use a top-tier model vs. a less intensive one.

Profilbild von RowdyGoose
RowdyGoosevor 2 Tagen

Semianalysis has 0 credibility any more

Profilbild von StockTake
StockTakevor 2 Tagen

Feels like lazy analysis. Open source taking over simpler tasks doesn’t necessarily weaken the frontier labs. It could massively expand the market while pushing frontier models to problems that weren’t economically feasible to solve before. The bigger question is whether the labs can maintain pricing power as intelligence gets cheaper. That’s what will determine who captures the value, not how many hypothetical PhDs the economy can absorb.

Profilbild von Nacho
Nachovor 2 Tagen

Once an app lets users pick, routine work drifts to the cheaper model, and the frontier labs have to earn their price on the hardest tasks

Profilbild von BullBear.News
BullBear.Newsvor 2 Tagen

the transition point is when local inference cost drops below the coordination overhead of external APIs

Profilbild von Sterling
Sterlingvor 2 Tagen

Unless you're a Big Tech SWE, the cost effectiveness of open source tokens is really really hard to ignore.

Profilbild von Nessuno
Nessunovor 2 Tagen

See Wallis/North 1986, or our book forthcoming @PalgraveEcon for an unqualified, No.

Profilbild von Tim
Timvor 2 Tagen

Competition at any level (state/country, corporation, individual) ensures that there is a market for ever higher levels of intelligence. You want to be smarter than your competitor.

Ähnliche Videos

Baseten Head of AI Model Training Charlie O'Neill says the future is many specialized LLMs dedicated to specific tasks, with bigger labs deployed on the frontiers of areas like science and math: "People are thinking about intelligence capabilities in the wrong way. People are thinking about intelligence relativistically. They say, 'OK, the open-source gap is like 6 months behind closed-source, and GLM 5.3 is as good as Opus 4.8,' or whatever." "The best way to think about what models can do for you, and for the world, is in an absolute sense." "So for any given task that you want to do with an LLM, there's some intelligence threshold where below that you can't do the task, and above that you have very diminishing returns to more intelligence on the task." "So when you think about it that way, the game of LLMs over the last 5 years has been, 'OK, we have these things we want to do with them. Closed source hits it first... but open-source can eventually do that task. And then for many reasons, once you have the base level of intelligence required to do it, you probably do want to swap to open-source." "It's not really about the [frontier lab] God model being better. Like, if I'm filing a tax return, there is a limit to how much intelligence I need to do that particular thing." "So I think the world is going to look like — frontier closed-source labs are going to continue to push the frontier. You do want to use the most intelligent model. You have very inelastic demand for intelligence when you're doing frontier science or frontier math." "But for a lot of the economically valuable things, it looks a lot like, 'I'm a Cursor, or I'm one of these big companies who are realizing I can't just be a wrapper anymore. I've been through the life cycle of building a product that people love. And I should be using that information to make my model better at the things that I care about, and not at anything else.'"

TBPN

50,531 Aufrufe • vor 1 Monat

David Sacks says companies are trapped paying OpenAI & Anthropic because they can't figure out how to use open source models "I think enterprise CTOs would like to shift their token consumption to cheaper models for the obvious reason that it would be more efficient. They are seeing compute costs or token costs skyrocket right now, so everyone's trying to figure this out." "You also have the AI sovereignty issue that Alex Karp talked about. They're worried about giving up the secret sauce or the alpha in their business to a frontier lab that may one day be competing with them. "The problem is, I think in most cases, they don't have the technical ability to do it. Coinbase figured out how to do it. DoorDash figured out how to do it. They built a token routing system that allows them to send frontier tasks to frontier models and non frontier tasks to more mundane models. But I don't think your average enterprise has the technical capability to do that." "This is why the share of wallet of closed models, it actually increased. I think that open source went from 19% last year to 11% this year. So open source as a share of enterprise spending is actually decreasing." "I don't think that means usage is decreasing. I think usage is skyrocketing. It also may be the case that because the whole point of using an open model is you just pay for the compute costs, you don't have to pay a lab, so it may be that it's hard to measure that usage in terms of spend." "But nonetheless, anyone who's saying that these closed models are going to lose or are somehow losing, you're just not seeing it in the data."

dnap

141,614 Aufrufe • vor 3 Monaten

learned a lot from this conversation with Simon Mo and Matt Bornstein. biggest takeaways for me: -there are a lot of reasons why we should like open-weight models. a lot of these arguments stop at handwavy things like "what if the labs stop releasing frontier models to the public" or "it's lower cost." but simon's position as lead maintainer of vLLM and CEO of Inferact give him authority to talk about some of the other, more interesting and concrete reasons to pay attention to open-weight models, namely that they allow end-users to calibrate latency / other performance metrics with way more customizability than what any of the frontier closed-source labs offer (and without the fear that your job might be met with a refusal at some random point where you're deep in a 2 hour job) -re: the above point...for this reason, a lot of US companies (inferact included!) choose to use open-weight models over their closed-source alternatives. this also isn't limited to internal workloads / research - on a recent a16z podcast the team at Decagon spoke about how something like 90% of their customer service ai agents run on open-weight models that they've fine-tuned. -we should really appreciate how many companies/teams came out researchers fascinated by the wave of very small open-weight models that were being distilled from e.g. gpt-3.5 and earlier models in 2022/2023 (prior to the release of chatGPT!). these small models motivated the development of pagedattention, which then led to vlmm/inferact (at other layers of the stack with similar origin stories, you can look at teams like openrouter or ollama). in other words, we have open-weight models to thank for a bunch of the orchestration infra we now rely on. i think yet another, indirect, way we can point to open-source/weight infra pushing the frontier forward. anyway, a lot more in this convo, it was a lot of fun!

Elena

12,922 Aufrufe • vor 2 Monaten