Loading video...

Video Failed to Load

Go Home

OpenAI's New Agent Stack: Superhuman Computer Use, Dots, Decisions API, UltraFast, & the AI Cloud OpenAI’s Ari Weinstein and Nikunj Handa explain why Computer Use is already becoming faster than humans at real-world tasks, how Dots gives agents their own cloud computers, why better models can now debug and...

24,541 views • 2 days ago •via X (Twitter)

6 Comments

Kizuno18's profile picture
Kizuno182 days ago

@OpenAI @AriX @nikunjhanda superhuman computer use speed only matters if action execution is coupled to accessibility tree state: raw coordinate clicks miss dynamic elements under fast renders. in our autonomous pipeline, executing actions against accessibility nodes keeps computer use 100% synchronized

Forge Orbital Inc's profile picture
Forge Orbital Inc2 days ago

@OpenAI @AriX @nikunjhanda Agent stacks keep getting more capable. The question regulated buyers ask next: what was this agent allowed to do, and can we prove it later? We built Forge for that: signed, replayable records on your own hardware. Would love to talk it through on the show.

Nikhil's profile picture
Nikhil2 days ago

@OpenAI @AriX @nikunjhanda Dots versus Grok Bot versus Metamuse. Is gonna win?

EKOS _ AGI 🇮🇷's profile picture
EKOS _ AGI 🇮🇷2 days ago

The interesting part isn’t just that agents can now operate faster and recover from failures. It’s the amount of execution history this creates. Decisions, tool calls, failures, recoveries and verified outcomes can become a compounding advantage if they’re preserved as reusable knowledge. Otherwise every faster agent is still starting from zero. Do you agree? @sama

Oleg's profile picture
Oleg2 days ago

@OpenAI @AriX @nikunjhanda my agents drive a real chrome. half their rules exist to slow them down

TobiasDev 🇦🇷's profile picture
TobiasDev 🇦🇷2 days ago

@OpenAI @AriX @nikunjhanda computer use faster than humans plus agents with their own cloud, the stack talk is getting concrete

Related Videos

AI INTERVIEW: OPENAI'S SECRET WEAPON AI agents are no longer just hype—they're here to revolutionize automation, Web3, and beyond. SwarmNode.ai is building a serverless AI agent platform for scalability, efficiency, and real-world impact. In this exclusive interview, he reveals how AI swarms can outperform single models, why OpenAI’s Operator is just the beginning, and how crypto is fueling AI innovation. Plus, he breaks down DeepSeek’s game-changing AI breakthrough, the future of agent monetization, and why serverless AI could be the next frontier in automation. 01:37 – From Engineering to AI: The journey into artificial intelligence. 02:43 – The GPT-3 Moment: How OpenAI’s tech pulled him in. 04:10 – AI’s Biggest Challenge: Why real-world use cases lag behind. 05:05 – OpenAI’s Operator: Why it’s “rudimentary” (for now). 06:25 – Crypto & AI: How tokens help bootstrap AI startups. 08:15 – Can You Bootstrap a Startup with a Token? The trade-offs. 09:56 – 90% of AI Token Holders Don’t Use the Product—Does It Matter? 11:18 – What is SwarmNode?: AI agents, hosted serverlessly. 14:23 – AI Swarms: Why multiple agents outperform single models. 16:08 – What is a Swarm? A simple definition of collaborative AI. 17:32 – “How Can I Make Money with AI?”: Real-world use cases. 18:41 – AI Bounties: Hiring devs to build your custom agent. 20:50 – The Future of AI Marketplaces: Monetizing pre-built agents. 23:15 – DeepSeek’s Disruption: Why it’s good news for AI. 24:46 – Is SwarmNode Compatible with DeepSeek? How it integrates. 26:17 – SwarmNode vs. AI Launchpads: What makes it different? 27:42 – Why Serverless Matters: Cost savings & efficiency. 29:53 – AI Agents in the Real World: Booking flights, managing workflows, and more. 31:11 – Building SwarmNode for Developers: Why it started as a personal project. 32:27 – Explosive Growth: 200,000 AI agent executions in 5 weeks. 34:41 – Why SwarmNode Agents Aren’t Visible on 𝕏 Yet. 36:46 – Startup Hiring Lessons: Finding top AI talent. 39:15 – Why SwarmNode is Built in Python (and What’s Next). 40:32 – Scaling AI Workloads: Handling traffic surges. 41:42 – AWS & Cost Challenges: The biggest monetization hurdle. 42:58 – 2025: The Year of Mass AI Adoption. 45:22 – Should We Be Worried About AI’s Rapid Growth? 46:46 – The Most Underrated AI Tools Right Now. 47:34 – What’s Next for SwarmNode?: Making AI accessible to everyone.

Mario Nawfal

338,311 views • 1 year ago

In the future, you’ll be able to accomplish a goal by just giving Claude an outcome and a budget. That’s the direction Anthropic is building in with its new Managed Agents features, announced at this week’s Code with Claude developer event. The basic idea: Claude, wrapped in a computer in the cloud, that you can spin up, scale, and manage as needed. Anthropic is taking on the infrastructure that kills most agent products, and making sure that it scales to meet the needs of agents running 24/7. On this week’s AI & I from Every 📧, I talk with Angela Jiang (Angela Jiang), head of product for the Claude platform, and Katelyn Lesse (Katelyn Lesse), head of engineering for the Claude platform, about what Anthropic is building and what it takes to make agents reliable in production. We get into: - Why the "build a generic harness, hot-swap any model behind it" playbook is already outdated. Angela points to eval data on Memory where the same task across different harnesses performed drastically differently. - The infrastructure wall every team hits in production—and why Katelyn thinks “my sandbox died and took the agent with it” is the real reason internal agents don't ship. - Why Anthropic is so bullish on using file systems and skills within Claude, including Angela's argument that those early design choices can compound for years. This is a must-watch for anyone trying to take an agent past the demo and into production. Watch below! Timestamps: How the Claude platform evolved from API to agents: 00:01:48 The primitives that make up Claude Managed Agents: 00:04:09 Why the harness and the model are becoming a single unit: 00:10:37 The infrastructure wall that kills most agent projects in production: 00:18:49 Why team agents need a different shape than individual productivity tools: 00:24:49 How Anthropic's legal team uses an agent to review marketing copy: 00:26:36 Using multi-agent orchestration for advisor strategies, adversarial pairs, and swarms: 00:34:24 How to measure agent success with outcome and budget as the end state: 00:35:50 What the platform looks like a year from now, when Claude writes its own harness: 00:39:11

Dan Shipper

66,871 views • 4 months ago