
Nerd Snipe
@NerdSnipePod • 3,021 subscribers
Nerd Snipe with @theo and @davis7.
Videos

Opus 5.5 is so good it finally got us to agree on something. 0:00 Intro 3:18 GPT-6 Sol & Luna 14:58 Grok 4.7 38:47 Jev explained 52:48 Why model routing fails 1:05:06 Opus 5.5 1:26:17 Pricing & usage limits 1:34:23 Building games with Opus 1:41:10 Porting TypeScript to Rust 1:48:57 Agent workflows & testing
Nerd Snipe84,620 görüntüleme • 5 gün önce

Mom and Dad are fighting again, and this time, it’s over Astra. Ben and Theo fight the entire podcast, let’s see about what. 00:00 Meet Astra 04:03 Best Model Ever, With Catches 06:33 Reasoning and 3D Benchmarks 20:00 Astra Rebuilds 30:04 Instruction-Following Problems 44:47 The Uncommitted Fix Debate 01:09:37 The PR Babysitting Failure 01:22:41 Why Fable Still Wins 01:30:41 Multimodal and Computer Use 01:39:29 Astra vs. Fable Fleet Data
Nerd Snipe154,054 görüntüleme • 1 ay önce

OpenAI vs. Anthropic, a mysterious 300B model, and the definitive AI model tier list (drinking game). What could go wrong? 0:00 Intro 3:30 OpenAI vs. Anthropic 14:44 Anthropic’s delayed models 20:53 Kimi K3 and open weights 30:08 The Alpha stealth model 48:26 Model tier list begins 1:09:53 GPT-5.6 Luna 1:13:09 Opus and Sonnet 1:20:25 Gemini models 1:31:15 DeepSeek and local models 1:59:58 Fable vs. Sol 2:28:46 Final rankings
Nerd Snipe117,810 görüntüleme • 1 ay önce

Codex went from basically unlimited to burning your entire limit in one prompt. We talked about how OpenAI cooked the $200 plan and quietly killed the Codex app. 0:08:46 Intro 0:10:15 Muse Spark 1.1 0:13:13 GPT-Live-1 0:22:03 Grok 4.5 0:35:58 Windsurf & RIP Codex 0:51:38 Apple Sues OpenAI 0:58:05 Elon vs Sam 1:13:24 GPT-5.6 Sol & Limits 1:35:19 Subagents & Agent Workflows 1:47:42 Claude Code + Sol
Nerd Snipe94,862 görüntüleme • 2 ay önce

We've spent six figures in tokens testing OpenAI's 5.6 Sol model to see whether if its better than Fable and what OpenAI have done to make it even better than 5.5. 02:38 - Intro 04:10 - What Is this New Model 07:17 - First Impressions 16:38 - The Code-Quality gap 26:38 - OpenAI's Naming Problem 29:38 - The Complicated Release 32:08 - Loops & Subagents 38:28 - macOS is killing your machine 46:28 - Claude Code vs Codex Subagent UX 1:03:19 - The Pottery Analogy 1:09:33 - cispolicyd CPU Usage
Nerd Snipe84,148 görüntüleme • 2 ay önce

Fable was supposed to give us 14 days and instead we got 3 before the export ban. Now it's back at half the rate limits. 00:00 FDE revolution 10:29 AI Engineer Conference 15:03 AI skills & workflows 24:52 Anthropic vs Alibaba 32:19 OpenAI government stake 39:25 Fable returns 47:51 Fable limits & capacity 55:19 Fable pricing & reasoning 01:02:09 Fable safeguards 01:11:10 Sonnet 5 vs GLM
Nerd Snipe65,914 görüntüleme • 3 ay önce

It seems we were wrong about Opus 5... 00:00 Intro 01:23 T3 Code inbox sidebar 13:29 Kimi K3 14:39 K3 cost and speed 23:17 Hugging Face hack 26:13 Open model ban 31:04 Breaking containment 36:03 The Dean post 42:32 Distilled models 51:46 Opus 5 1:04:02 3D and games 1:11:16 Model picking
Nerd Snipe34,186 görüntüleme • 2 ay önce

Anthropic Doesn't Think You Can Be Trusted, China is closing the gap, SpaceXAI Leaps Ahead, and DEF CON 0:00 Intro 2:56 Claude Watermarks 15:04 Meta + Muse 48:48 GLM-5.3 55:15 Qwen + DeepSeek 1:08:48 Grok 4.6 + Bot 1:20:15 Gavin vs Dario 1:43:44 DEFCON 2:15:35 Viewer Q&A
Nerd Snipe23,780 görüntüleme • 1 ay önce
Daha fazla içerik yok.