Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Everything Gets Rebuilt: my conversation with Harrison Chase, CEO of LangChain about agent harnesses, evals, runtimes, sandboxes, MCP and the future of the agent stack 00:00 Intro - meet Harrison Chase - at the Chase Center for the Daytona Compute conference 01:32 What changed in agents over the last...

35,253 Aufrufe • vor 5 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

How to build long-horizon AI agents: behavior specs, ontologies, process supervision - my conversation with Mitchell Troyanovsky, co-founder of Basis 01:09 Why Everyone at Basis Was Whispering to AI when Stephanie Palazzolo walked in 04:12 Accounting as "an Intelligence Over the Economy" 06:11 What Makes an Agent Truly Long-Horizon 08:24 Inside an Autonomous, Multi-Day Tax Return 10:19 Agents That Hand Off Like Senior Engineers 11:17 A Brief History of Agents: From ReAct to Today 12:33 Why LLMs Have No Long-Term Memory 14:13 Why AutoGPT Didn't Live Up to Its Promise 15:51 The Three Breakthroughs: Opus 3, o1, o3 17:07 Why Reasoning Models Unlocked Agents 18:23 "Let's Verify Step by Step": The Road Not Taken 20:32 Pushing Back on the METR Chart 22:09 Why Coding Agents Won First 25:14 Why Real-World Agents Are Harder 26:55 How Accountants Verify Non-Deterministic Work 29:18 You Can't Scale Tax Returns Like Math 33:16 100 Evals Pass - So What? 35:53 Right Answer, Wrong Process 36:37 Behavior Specs, Explained 39:58 How Specific Should Behaviors Be? 42:18 Context Is Runtime Training Data 44:21 Who Judges the Judge? 46:45 The Move 37 Objection 50:02 The Magic Box Mental Model 52:41 "Nothing Has Changed Since o3" 54:56 Open-Sourcing Behavior Specs with Ankur Goyal Braintrust 59:45 Ontologies: A World for Agents to Live In 01:04:20 Documentation as Codebase 01:06:33 Why the Founding Fathers Were Context Engineers 01:09:05 Onboarding 300 Brilliant Alien Employees 01:11:10 Self-Improving Agent Systems 01:12:50 The Context Mistake Agent Builders Make 01:14:29 RL on Behavior Adherence 01:17:01 Will the Bitter Lesson Swallow the Harness 01:18:46 "Technical Moats Are Not Real Moats" 01:21:03 Advice for AI Builders

Matt Turck

20,898 Aufrufe • vor 25 Tagen

AI INTERVIEW: OPENAI'S SECRET WEAPON AI agents are no longer just hype—they're here to revolutionize automation, Web3, and beyond. SwarmNode.ai is building a serverless AI agent platform for scalability, efficiency, and real-world impact. In this exclusive interview, he reveals how AI swarms can outperform single models, why OpenAI’s Operator is just the beginning, and how crypto is fueling AI innovation. Plus, he breaks down DeepSeek’s game-changing AI breakthrough, the future of agent monetization, and why serverless AI could be the next frontier in automation. 01:37 – From Engineering to AI: The journey into artificial intelligence. 02:43 – The GPT-3 Moment: How OpenAI’s tech pulled him in. 04:10 – AI’s Biggest Challenge: Why real-world use cases lag behind. 05:05 – OpenAI’s Operator: Why it’s “rudimentary” (for now). 06:25 – Crypto & AI: How tokens help bootstrap AI startups. 08:15 – Can You Bootstrap a Startup with a Token? The trade-offs. 09:56 – 90% of AI Token Holders Don’t Use the Product—Does It Matter? 11:18 – What is SwarmNode?: AI agents, hosted serverlessly. 14:23 – AI Swarms: Why multiple agents outperform single models. 16:08 – What is a Swarm? A simple definition of collaborative AI. 17:32 – “How Can I Make Money with AI?”: Real-world use cases. 18:41 – AI Bounties: Hiring devs to build your custom agent. 20:50 – The Future of AI Marketplaces: Monetizing pre-built agents. 23:15 – DeepSeek’s Disruption: Why it’s good news for AI. 24:46 – Is SwarmNode Compatible with DeepSeek? How it integrates. 26:17 – SwarmNode vs. AI Launchpads: What makes it different? 27:42 – Why Serverless Matters: Cost savings & efficiency. 29:53 – AI Agents in the Real World: Booking flights, managing workflows, and more. 31:11 – Building SwarmNode for Developers: Why it started as a personal project. 32:27 – Explosive Growth: 200,000 AI agent executions in 5 weeks. 34:41 – Why SwarmNode Agents Aren’t Visible on 𝕏 Yet. 36:46 – Startup Hiring Lessons: Finding top AI talent. 39:15 – Why SwarmNode is Built in Python (and What’s Next). 40:32 – Scaling AI Workloads: Handling traffic surges. 41:42 – AWS & Cost Challenges: The biggest monetization hurdle. 42:58 – 2025: The Year of Mass AI Adoption. 45:22 – Should We Be Worried About AI’s Rapid Growth? 46:46 – The Most Underrated AI Tools Right Now. 47:34 – What’s Next for SwarmNode?: Making AI accessible to everyone.

Mario Nawfal

338,265 Aufrufe • vor 1 Jahr

An agent is three things: a harness, a model, and context. If you're serious about owning your intelligence, you probably want to own all three. LangChain founder Harrison Chase joined us at our Sequoia Capital Own Your Intelligence to talk about the piece that often gets the least attention: the harness. He offers a clear heuristic for when to build your own. The more out of distribution you are from what the models were trained on, the more you'll want to customize. And good technical content on how to actually measure performance with evals and langsmith. 00:00 Introduction 00:58 The three parts of an agent: harness, model, context 02:12 What a harness actually does 03:25 Customizing the core loop with middleware 04:41 Sandboxes, file systems, sub-agents, summarization 05:47 Cognitive architectures — and when you still need them 07:03 Build your own harness or use off the shelf? 08:24 In-distribution vs. out-of-distribution: the file-editing example 09:39 Why evals define what "good" means in an organization 11:04 Harbor: what an eval task actually looks like 12:11 Comparing harnesses and models on accuracy, latency, and cost 13:20 Why observability is underrated — it's usually the context 14:34 The data flywheel: traces → curation → experiments 15:42 Getting feedback through UX design and online evaluators 16:51 Demo: LangSmith Engine 19:23 Q&A: Running Engine on Engine, and "codex-ification" 20:44 Q&A: Will harnesses converge or diverge?

Sonya Huang 🐥

75,702 Aufrufe • vor 18 Tagen