Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Are Chinese open models being oversold as reaching the frontier? We’re back with another vibe check episode with Ari Morcos and Rob Toews. These are a ton of fun to record and this was a particularly meaty one given everything going on. We hit on: ▪️Did Chinese open models...

14,496 Aufrufe • vor 1 Monat •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

"As a proprietary model builder, you're kind of starting to get squeezed in.” That's Igor Babuschkin’s take on where the biggest AI labs stand today. He explains how you need the models to get way better to keep margin, but at some level they may be too sensitive to release. This week, I sat down with Igor on Unsupervised Learning. It was a fascinating conversation with someone who has real perspective on the questions everyone in AI is asking right now. From DeepMind's StarCraft project to early reasoning work at OpenAI to co-founding xAI, Igor has had a front-row seat to nearly every major AI breakthrough. He's now the co-founder of River AI, building individualized, locally-run AI models. We discuss: ▪️Why proprietary model labs might be in trouble ▪️What it's like working with Elon ▪️Building Colossus in 120 days ▪️Should enterprises train their own models ▪️What’s left for humans as models get better ▪️Why he left xAI to bet on personal local AI instead ▪️The three bets River is taking ▪️What's stopping AI from moving beyond coding ▪️Reflections on the rapid pace of the past years 0:00 Intro 1:17 Writing Fiction on Where AI Is Headed 4:46 Cracking Agents Beyond Coding 10:29 Why Igor Left to Start River 12:22 River's Three Big Bets 18:06 Weights vs. Memory: The Personalization Debate 22:04 Should Enterprises Train Their Own Models? 25:10 Are Proprietary Labs Losing Their Edge? 32:16 The China Open-Source Problem 44:19 The Elon Call That Started xAI 50:18 Thoughts on Cursor Acquisition 52:16 What's Actually Bottlenecking AI 56:55 Humans, Machines, and Staying Relevant 1:01:29 Igor's Odds This All Goes Well YouTube: Spotify: Apple:

Jacob Effron

167,826 Aufrufe • vor 1 Monat

🦙 ollama is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan (Jeffrey Morgan) a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of Lightcone Podcast, Jeff joins Garry Tan, Jared Friedman, Diana, and Harj Taggar to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. 00:43 — The Shift to Open Models 03:03 — How AI Agents Are Driving Token Usage 05:31 — Are Open Models Catching Up? 08:26 — What Happens When a New Model Launches 11:31 — Ollama as an Operating System for AI 14:05 — The New Opportunities Above the Model Layer 18:19 — Why 80–90% of Enterprise Tokens Could Be Open 20:57 — The Future Is Local and Cloud 26:40 — Why AI Is Coming Back to Your Computer 28:56 — The Coming Era of Unlimited Tokens 32:30 — Do We Still Need a “God Model”? 33:41 — Open Models and Geopolitics 36:14 — The Origins of Ollama 40:36 — Two Years Lost in the Wilderness 42:39 — The Pivot That Changed Everything 47:02 — How Ollama Found a Business Model 49:43 — Why Second-Time Founders Did YC

Y Combinator

306,787 Aufrufe • vor 7 Tagen

Thanksgiving-week treat: an epic conversation on Frontier AI with Lukasz Kaiser -co-author of “Attention Is All You Need” (Transformers) and leading research scientist at OpenAI working on GPT-5.1-era reasoning models. 00:00 – Cold open and intro 01:29 – “AI slowdown” vs a wild week of new frontier models 08:03 – Low-hanging fruit, infra, RL training and better data 11:39 – What is a reasoning model, in plain language 17:02 – Chain-of-thought and training the thinking process with RL 21:39 – Łukasz’s path: from logic and France to Google and Kurzweil 24:20 – Inside the Transformer story and what “attention” really means 28:42 – From Google Brain to OpenAI: culture, scale and GPUs 32:49 – What’s next for pre-training, GPUs and distillation 37:29 – Can we still understand these models? Circuits, sparsity and black boxes 39:42 – GPT-4 → GPT-5 → GPT-5.1: what actually changed 42:40 – Post-training, safety and teaching GPT-5.1 different tones 46:16 – How long should GPT-5.1 think? Reasoning tokens and jagged abilities 47:43 – The five-year-old’s dot puzzle that still breaks frontier models 52:22 – Generalization, child-like learning and whether reasoning is enough 53:48 – Beyond Transformers: ARC, LeCun’s ideas and multimodal bottlenecks 56:10 – GPT-5.1 Codex Max, long-running agents and compaction 1:00:06 – Will foundation models eat most apps? The translation analogy and trust 1:02:34 – What still needs to be solved, and where AI might go next

Matt Turck

168,007 Aufrufe • vor 9 Monaten

Been using Nous Research Hermes Agent for media editing and it completely replaced my workflow. You know the drill : - Trim a video - open CapCut. - Convert to mp3 - another site. - Remove audio - another one. - Make a gif - yet another one. - more tasks, more different tools minutes of uploading and downloading.. Now you can do all of that just by chatting like asking a designer friend to edit your media and getting it back instantly :) I built a media toolset for Hermes that handles all of this natively. It registers as a skill inside the agent - Hermes sees your video, understands what you're asking, takes the right operations and use the best skills and give you what you really wanted perfectly ▪️ Trim any segment down to the second ▪️ Speed up or slow down with pitch corrected audio ▪️ Compress without destroying quality ▪️ Create a gif on defined times range ▪️ Cut the video and rotate it and apply instagram ▪️ Convert between any format - mp4, mp3, ogg, webm ▪️ Chain edits: "trim 0:30 1:00 and speed up 1.5x and compress" ▪️ Burn .srt subtitles directly into the video ▪️ Add or remove text watermarks with custom position ▪️ Generate thumbnails from any timestamp ▪️ Fade in/out with custom duration ▪️ ASCII video art with multiple character styles and many more with customizable skills Works the same on Telegram, Discord, WhatsApp etc. This is what agents should be doing and Hermes does.

ogichain

11,291 Aufrufe • vor 6 Monaten

Inside Nemotron and NVIDIA's AI lab: my conversation with Bryan Catanzaro (Bryan Catanzaro). NVIDIA is a chip company. So why does it put hundreds of researchers on building AI models - and then give them away for free? We go deep into the Nemotron models, what it takes to build a top AI lab, and the future of frontier AI. 01:33 - Is open source AI catching the frontier? 05:29 - Do closed labs blocking distillation slow open source down? 07:42 - Is the US falling behind China? 10:30 - Why companies actually choose open models 12:39 - A "crazy" 2008 bet: machine learning on GPUs 15:33 - Working with Andrew Ng and Dario Amodei at Baidu 17:41 - Coming back to NVIDIA: DLSS and the birth of Megatron 21:55 - The real reason NVIDIA builds its own models 24:28 - Is Moore's Law really dead? 33:37 - The Nemotron family: Nano, Super, Ultra 35:09 - Built for agents: why NVIDIA bets on speed 36:02 - How you train a 550B model in 4 bits 39:25 - Hybrid Mamba-Transformer, explained simply 42:31 - Mixture of experts, and why NVIDIA built NVL72 around it 47:26 - Why a 1-million-token context window matters 49:26 - Multi-token prediction: how the model predicts 5 tokens at once 52:47 - Multi-teacher distillation: teaching one model from many 58:01 - Where reinforcement learning goes next 01:00:16 - Inside NVIDIA's research org: "the mission is the boss" 01:04:03 - How NVIDIA decides who gets the GPUs 01:10:53 - Why NVIDIA still feels entrepreneurial after 33 years 01:12:58 - Why Bryan doesn't believe in the singularity 01:17:50 - The AI backlash 01:19:18 - The controversial case: open AI is safer than closed

Matt Turck

56,954 Aufrufe • vor 2 Monaten