Загрузка видео...

Не удалось загрузить видео

На главную

.OpenRouter co-founder Alex Atallah, in his first podcast since Stripe acquired the company, joins Replit ⠕ co-founder Amjad Masad and a16z's Erik Torenberg on why the future of AI is independence and specialization. In this conversation, Alex walks through how the Stripe deal unfolded, why he wasn't originally looking...

546,007 просмотров • 4 дней назад •via X (Twitter)

Комментарии: 21

Фото профиля Tyler Edwards
Tyler Edwards3 дней назад

some numbers from our own testing. contract clauses, 4,000+ questions across 100+ contracts. base model 80.4 F1, frontier flagship 82.0, the same base model trained on the task 91.6. quoting the clause word for word: frontier 8.5%, trained 59.6%. so model size wasn't what moved it, the task data was

Фото профиля Anderson
Anderson3 дней назад

@OpenRouter @Replit specialization is just fragmentation with funding

Фото профиля Acezhang
Acezhang4 дней назад

@OpenRouter @Replit i'm with alex, partly for a boring reason: 10 specialized agents means 10 small permission scopes. one superagent with access to everything is one misread instruction away from touching all of it

Фото профиля Jean Pierre · fireply.ai
Jean Pierre · fireply.ai4 дней назад

@OpenRouter @Replit the lab that supplies you can become your competitor. like a baker who also opens a bistro across the street, mais bien sûr

Фото профиля Sterling
Sterling4 дней назад

@OpenRouter @Replit This is why I built Haldir to be model agnostic. Agents should be customized and so should the guardrails you can put on them with Haldir

Фото профиля Thomas Jankowski
Thomas Jankowski4 дней назад

Routing makes the model a commodity, and the router sits on the one thing that never commoditizes: the invoice. If every call already flows through a single key, the markup, the credits float and the spend data come with it. That's a payments business, not an inference one. Stripe doesn't need the routing to be clever. It needs the bill to be the switching cost, since swapping a model takes an afternoon and unwinding consolidated billing, budgets and chargebacks takes a quarter. Does independence survive once the router is owned by the party that settles the money, doesn't it?

Фото профиля kennek
kennek4 дней назад

@OpenRouter @Replit specialization is correct and also how my company AI search died. we specialized so hard the only person who knew which tool to open was me. everyone else slacked Dave.

Фото профиля Chloee💝
Chloee💝3 дней назад

@OpenRouter @Replit the 10 agents thing makes sense until you realize one of them will quietly become the glue and then youre basically back to one superagent with extra steps

Фото профиля NIKKI
NIKKI4 дней назад

@OpenRouter @Replit Specialized agents make sense, but cross-domain joins are hard to give up.

Фото профиля Tyzo
Tyzo4 дней назад

@OpenRouter @Replit Independence and specialization in AI sound really promising

Фото профиля Desmond Lim
Desmond Lim4 дней назад

@OpenRouter @Replit Specialization makes sense as models become more capable but less differentiated. I wonder if the enduring layer will be the model itself—or the routing, context and workflow that consistently chooses the right model.

Фото профиля ⋆˙ sofi ✧
⋆˙ sofi ✧4 дней назад

@OpenRouter @Replit independence and specialization is a nice way of saying everyone's about to build their own thing

Фото профиля Abdul Q.
Abdul Q.3 дней назад

@OpenRouter @Replit Agency version of "everyone's building the same agent." The gap is still the brief, the Webflow craft, and knowing which page the client actually meant.

Фото профиля Rishi Dadhania
Rishi Dadhania3 дней назад

@OpenRouter @Replit @grok Why all are getting excited over this what tldr is it giving and what the person is telling about that would happen in Future Explain it to me and people who also wants to understand

Фото профиля Aakash
Aakash4 дней назад

@OpenRouter @Replit ten chiefs, who's the staff?

Фото профиля David Arnal
David Arnal4 дней назад

@OpenRouter @Replit Specialization feels more compelling when cost and model choice are both moving targets, especially for routine tasks that don't need frontier models.

Фото профиля Y. Kh
Y. Kh3 дней назад

Both can be true actually. I'm a believe too that there is no way that the the future of work is managing 1000 agents. Every agent user knows that above 5 simultaneous agents, it becomes just impossible to keep track. So the future will probably be a single entry point, 1 agent that you talk too. However, it does not mean that this agent will do everything. We'll probably have many specialized agents underneath, called by this single agent. It's what is already happening at some level.

Фото профиля Graham Holloway
Graham Holloway4 дней назад

@OpenRouter @Replit AI model independence matters as enterprises balance cost, performance, and lock-in. I make room for these infrastructure debates here; @_OhHeyZiggy is another frequent stop.

Фото профиля async
async4 дней назад

@OpenRouter @Replit It’s been so nice not seeing Amjad anywhere for the past 12 months. Not all good things are meant to last it seems.

Фото профиля Raven
Raven4 дней назад

@OpenRouter @Replit i'd rather be one of ten agents than the bot blamed for every department's weird little crisis

Фото профиля 泰坦链 | DeFi
泰坦链 | DeFi4 дней назад

@OpenRouter @Replit 10 agents is the move

Похожие видео

Baseten’s CEO on why every company will want to own its AI / Tuhin Srivastava on the rise of AI agents, the demand for inference, and building a new kind of hyperscaler. Lately, I’ve been spending a lot of time thinking about the AI inference market and how big it could get. Agents like Muse, Instinct, Town, and Grok Bot aggressively use browsers, run their own computers, and burn through far more tokens than a traditional chatbot. As more people put them to work (Muse is number two in the App Store), the demand for inference, or the computing needed to run these models, should grow enormously. Tuhin and I discuss the rise of agents using browsers and virtual machines to get things done, and Baseten's recent acquisition to help power that shift. We also talk about the data center backlash, why companies are embracing Chinese open models, his plans for Baseten’s new research lab, and why he thinks inference becomes the only market left after AGI. Timestamps: 00:00 What Is AI Inference? 06:05 Competing With the Cloud Giants 09:24 Why Companies Want to Own Their AI 15:30 Building Baseten Before the AI Boom 23:45 DeepSeek and the Race for Open AI Models 29:32 Baseten’s Growth and Expansion 33:05 AI Agents and the Blaxel Acquisition 38:32 Data Centers and the AI Backlash 43:25 What Happens to Inference After AGI? 45:12 When AI Agents Become Customers Thanks to the show's premier sponsors: Atlassian, Granola, and Mercury.

Alex Heath

26,164 просмотров • 22 дней назад

Amjad Masad with Gagan Biyani and Erik Torenberg on HAA: The Horowitz Andreessen Academy The ideal education today is a place with room to meander, a space to go heads-down when you’re obsessed, alongside unconventional peers. Gagan sees it in every student he meets: the interesting thing they're building is happening outside class, unsupported. Amjad got really into chess and it evolved into a whole product line. The thesis behind HAA is side quests shouldn't be on the side. No grades, no tests. Trust is the design principle - Silicon Valley runs on it, so why not apply it to students? Empower ambitious young people to run with the ideas that have real gravity they can’t avoid. "Young people should be heretics… We respect the past, but we're unapologetically going to change it." 1:03 Why young people should be heretics 5:35 Amjad designs his ideal college 8:45 Side quests deserve more resources 10:55 Don't take a Python class 14:30 College vs. raising money at 18 17:15 Most teen startups are premature optimization 18:30 "I tried not to start Replit" 19:40 How a chess hobby became a Replit product 24:25 The self-driving company 26:30 Respect the past, change it anyway 28:50 What Oakland politics taught Gagan about trust 31:50 Why trust works better than rules 36:35 The positive-sum culture HAA wants to pass on 39:20 Speaking to the world vs. speaking to the in-group 41:05 How Amjad found his moral compass 44:00 What Amjad looks for in young talent YouTube: Amjad Masad Gagan Biyani 🏛 Erik Torenberg

a16z

167,988 просмотров • 14 дней назад

🦙 ollama is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan (Jeffrey Morgan) a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of Lightcone Podcast, Jeff joins Garry Tan, Jared Friedman, Diana, and Harj Taggar to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. 00:43 — The Shift to Open Models 03:03 — How AI Agents Are Driving Token Usage 05:31 — Are Open Models Catching Up? 08:26 — What Happens When a New Model Launches 11:31 — Ollama as an Operating System for AI 14:05 — The New Opportunities Above the Model Layer 18:19 — Why 80–90% of Enterprise Tokens Could Be Open 20:57 — The Future Is Local and Cloud 26:40 — Why AI Is Coming Back to Your Computer 28:56 — The Coming Era of Unlimited Tokens 32:30 — Do We Still Need a “God Model”? 33:41 — Open Models and Geopolitics 36:14 — The Origins of Ollama 40:36 — Two Years Lost in the Wilderness 42:39 — The Pivot That Changed Everything 47:02 — How Ollama Found a Business Model 49:43 — Why Second-Time Founders Did YC

Y Combinator

313,654 просмотров • 1 месяц назад

Inside Nemotron and NVIDIA's AI lab: my conversation with Bryan Catanzaro (Bryan Catanzaro). NVIDIA is a chip company. So why does it put hundreds of researchers on building AI models - and then give them away for free? We go deep into the Nemotron models, what it takes to build a top AI lab, and the future of frontier AI. 01:33 - Is open source AI catching the frontier? 05:29 - Do closed labs blocking distillation slow open source down? 07:42 - Is the US falling behind China? 10:30 - Why companies actually choose open models 12:39 - A "crazy" 2008 bet: machine learning on GPUs 15:33 - Working with Andrew Ng and Dario Amodei at Baidu 17:41 - Coming back to NVIDIA: DLSS and the birth of Megatron 21:55 - The real reason NVIDIA builds its own models 24:28 - Is Moore's Law really dead? 33:37 - The Nemotron family: Nano, Super, Ultra 35:09 - Built for agents: why NVIDIA bets on speed 36:02 - How you train a 550B model in 4 bits 39:25 - Hybrid Mamba-Transformer, explained simply 42:31 - Mixture of experts, and why NVIDIA built NVL72 around it 47:26 - Why a 1-million-token context window matters 49:26 - Multi-token prediction: how the model predicts 5 tokens at once 52:47 - Multi-teacher distillation: teaching one model from many 58:01 - Where reinforcement learning goes next 01:00:16 - Inside NVIDIA's research org: "the mission is the boss" 01:04:03 - How NVIDIA decides who gets the GPUs 01:10:53 - Why NVIDIA still feels entrepreneurial after 33 years 01:12:58 - Why Bryan doesn't believe in the singularity 01:17:50 - The AI backlash 01:19:18 - The controversial case: open AI is safer than closed

Matt Turck

56,954 просмотров • 3 месяцев назад