Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

How do we cache prompts with LLMs? Interview with Tanishq Singh. This is a real interview question from a big tech company, asked to a candidate in their technical interview round. The video explains the answer in roughly 6 minutes. 00:00 Question - Prompt Caching with LLMs 00:39 Exact...

86,167 görüntüleme • 5 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

AI has a trust problem. Verifiability is the solution. Our GM of AI Nima Vaziri sat down with a16z’s Ali Yahya and Dan Boneh of Stanford University to map the deepest fault lines in AI today. ☁️ Models we can’t trust ☁️ Current providers can censor, shut down, or shift rules overnight. Outsourced training hides backdoors. Even “open” weights don’t prove what’s actually running. Trust. Backdoors. Black boxes. The path forward is clear: 🔥 Verifiable evals 🔥 Verifiable inference 🔥 TEEs for hardware-backed integrity 🔥 Infra beyond single points of control 🔥 Blockchains as coordination layers for AI From “trust us” to “verify yourself.” That’s the shift. That’s the unlock. The frontier is here. The builders decide what comes next. Create and use AI that’s incentive aligned with you. Timestamps: 00:00:00 Introduction: AI & Crypto Intersection Overview 00:01:58 Four Major AI-Crypto Trends 00:02:44 AI Agents Need Financial Infrastructure 00:04:03 Proof of Humanity: Fighting AI-Generated Content 00:04:17 Decentralizing AI Infrastructure Networks 00:04:44 Synthetic Life: Autonomous AI Agents 00:06:20 Verifiable AI 00:10:16 Current Performance Numbers for AI Proofs 00:13:18 The Era of Experience in AI Learning 00:14:56 AI Agents Having Life of its Own 00:18:21 Algorithmic Fairness & Verifiable Models 00:23:18 Privacy in AI: Trusted Execution Environments 00:25:47 Economic Incentive for Open Weight Models 00:31:39 Attribution Problem: Who Gets Paid for AI Training? 00:35:52 Content Provenance & Authentication (C2PA) 00:48:03 AI Security: Finding Exploits & Vulnerabilities 00:54:53 Educational Applications: LLMs as Learning Partner 00:58:29 Reliance on LLMs and Cognitive Abilities 01:03:57 Content Providers’ Fear of LLM Training

EigenCloud

62,099 görüntüleme • 10 ay önce

I'm often asked for the best public example of AI evals done right for a real, production product. I finally have an answer. Teresa Torres shares how she shipped an AI interview coach, and used evals to rapidly squash bugs and improve the product. Teresa shows how she: 1. did error analysis FIRST to find real issues (instead of using generic metrics) 😍 2. used Jupyter notebooks to analyze errors 3. built custom annotation tools + custom widgets in notebooks 4. built a LLM-judge and assertions to test for specific errors 5. iterated through this feedback loop until it worked. 6. kept things simple the whole time It's also probably the best commercial for Jupyter notebooks you can imagine. 🥰 Chapter summary below. Link to YT in next thread 00:00:00 - Intro 00:01:45 - The Product: Building an AI Interview Coach 00:06:34 - The Problem: How Do I Know if My AI Coach is Any Good? 00:10:15 - Using Airtable for Traces and Annotation 00:12:15 - Discovering Jupyter Notebooks and Designing the First Evals 00:15:15 - Example Evals: LLM-as-Judge vs. Code-Based Assertions 00:21:00 - Learning Python with ChatGPT to Analyze Eval Results 00:31:00 - VS Code, Custom Tools, and an Eval Investigation Notebook 00:39:45 - Building a Custom Annotation Tool with Claude 00:41:00 - From Personal Project to Production App 00:46:02 - How Should PMs and Engineers Collaborate on AI Products? 00:55:45 - Q&A: Capturing Feedback and Annotations from End Users 00:58:11 - Q&A: Is a Technical Background Necessary to Build AI? 01:02:28 - Q&A: What's Next for Teresa? 01:03:13 - Q&A: Unpacking the Micro-Decisions of Building an AI App

Hamel Husain

51,376 görüntüleme • 11 ay önce