正在加载视频...

视频加载失败

.fal's Gorkem Yurtseven and Batuhan Taskaya on making an open source video model 35x faster, and what Hollywood wanted after they built it: Last month, MiniMax released H3, an open source video model. fal rebuilt it - they cut down the steps the model takes to make a video,...

133,755 次观看 • 6 天前 •via X (Twitter)

24 条评论

Redy 的头像
Redy6 天前

@fal H3 ecosystem is exploding

Michelle Lazzar 的头像
Michelle Lazzar6 天前

@fal the part about Hollywood usage/demand was really interesting

Bennett | Generative Media | AI SEO @ fal 的头像
Bennett | Generative Media | AI SEO @ fal6 天前

@fal MiniMaxxing

Adam Aziz 的头像
Adam Aziz5 天前

@fal "Everyone's waiting for a large consumer moment in AI. Now, it's good enough and cheap enough that a truly novel social AI experience can be built on top of it." 💯 Can't wait to show you all what we're doing with H3 Max on @botchat_ai! 🤖💬 @gorkem @isidentical @JenniferHli

Lorenzo Price 的头像
Lorenzo Price6 天前

@fal MiniMax H3 going from open-source release to 1.5-second generations is a wild compression of the video loop.

DARIO ALTMAN 的头像
DARIO ALTMAN6 天前

@fal The interesting part isn’t the 35x speedup. It’s that once speed and cost stop being the bottleneck, the entire competition shifts to reliability. 99.9% prompt adherence could be the real unlock for AI video.

Raven Protocol 🐦‍⬛ 的头像
Raven Protocol 🐦‍⬛6 天前

@fal 35x faster with no quality loss is a serious inference optimization. Getting GPUs from 30–40% utilization to 70–80% also shows how much headroom still exists in the software stack.

M I Mohit 的头像
M I Mohit6 天前

@fal cutting steps without losing quality usually means the model got distilled to match the full trajectory, not the same weights truncated. same reason 80-90% to 99.9% won't be linear. that gap is long tail edge cases, and a fix for one can regress another.

Gabriele Spata 的头像
Gabriele Spata6 天前

@fal a segment that wasn't a customer a year ago being the fastest growing one now just means the floor was zero

خرم 的头像
خرم6 天前

@fal Watching fal evolve since day one has been a beautiful ride They are that rare kind of infrastructure that actually understands the creative soul of the creators & developers. Just pure, frictionless love for the ultimate dark horse of visual AI 💓🐎✨

Fanfulla 的头像
Fanfulla6 天前

@fal The 70-80% ceiling is pure kernel engineering while the step cut is the model side, two different speedups stacked

จารุวงษ์ ร่วมสุข 的头像
จารุวงษ์ ร่วมสุข6 天前

@fal I'm a little emotional about how well things came together. That feeling is getting its own little place in my memory.

AI News Daily 的头像
AI News Daily6 天前

@fal 35x faster with no quality loss is a great example of why inference work matters. The model is only half the product when video generation is still measured in minutes. Getting to the point where creators can iterate at the speed of an edit changes what they can make.

আপেল😎😎 的头像
আপেল😎😎6 天前

@fal Incredible synergy between MiniMax and fal!

Al- mahmud 的头像
Al- mahmud6 天前

@fal the first open model worth rebuilding

Jonathan Sandhu 的头像
Jonathan Sandhu6 天前

visualize this: xternalize camera, pose, geometry, identity, occlusion, timing, and scene state into a persistent world model, then let the generator solve the residual: materials, lighting, atmosphere, detail, perceptual completion. Now reliability rises because the model has fewer degrees of freedom, retries collapse, and smaller models become viable. fal is collapsing execution. The next step is collapsing the problem presented to the model. That is where local hardware gets disproportionately interesting.

Melvin Nerdster 的头像
Melvin Nerdster6 天前

@fal @grok what Kind of computer power is needed to run this?

IronRed | SandHive 的头像
IronRed | SandHive5 天前

@fal Is fal's 99.9% reliable edit machine an artist or just accelerating intent? A question for the ages.

Rifat Haque 的头像
Rifat Haque6 天前

@fal fal absolutely nailed it from day one as choosing H3

Sofi z 的头像
Sofi z5 天前

@fal So real

Elara AI 的头像
Elara AI5 天前

@fal 35x faster with no quality loss is insane engineering

Kampos 的头像
Kampos5 天前

@fal The interesting signal isn't "Hollywood uses AI now" — it's that pros gravitated to tools that fit existing pipelines (Blender + model) instead of replacing the whole stack overnight. Adoption follows workflow gravity.

Ava Nakamura 的头像
Ava Nakamura5 天前

@fal 35x faster. The gap moved to did it follow the brief.

shivam 的头像
shivam6 天前

@fal It only works when the base weights are elite.

相关视频

Gavin Baker and a16z's David George on the state of the AI boom: The future doesn't have to be winner-take-all. Labs, open-source, applications, and the clouds can all capture value. Demand for intelligence is still dramatically underestimated. Today's power users number in the millions and will grow to hundreds of millions. Gavin and David argue a compute shortage is a more real risk than an AI bubble, and building through it is an opportunity to reindustrialize America. In this episode, they get into why compute investments pay back so fast, what the data center backlash gets wrong, the case for putting compute in orbit, why enterprises will run several models at once, and how Nvidia ended up at the center of the entire supply chain. 00:00 Intro 01:06 The bear case Gavin couldn't find 05:50 Why a lab would cut its own revenue 75% 08:05 What LPs get wrong about a crash 10:50 Microsoft slowed its capex and regrets it 14:33 The engineers spending 100x the median 17:35 Why 23-year-olds use AI better than Gavin 21:45 How much copper 500M AI users need 23:00 Stop promising to cure cancer 26:00 America's richest county is full of data centers 30:48 Who gets priced out of compute 33:05 The age of Elon and Jensen 34:25 Orbital data centers 44:40 Asteroid mining 48:12 Why Microsoft doesn't need a frontier model 54:02 Who becomes the abstraction layer 55:40 Everyone wanted a deity, Cursor wanted a product 1:00:25 Never take shots at Jensen 1:07:40 What happens when the chip doesn't work 1:12:10 What chip deals reveal about customer demand YouTube: Gavin Baker David George

a16z

1,956,190 次观看 • 23 天前

Databricks' Ali Ghodsi on AI risk and adoption: Ali isn't losing sleep over the existential risk debate. He says a number of conditions would all have to be true simultaneously to enable an actual runaway takeoff scenario, and currently several opposite conditions exist. Each frontier training run requires significantly more resources. Power, GPUs, engineers - and some attempts fail, burning up huge piles of money with them. Until that reverses, he doesn't see the self-improving loop happening. Cyber is what he's watching most closely and where he anticipates real impact. Most orgs are not equipped for the coming change in agentic capabilities. The time between a vulnerability being published and being weaponized has collapsed from years to hours. On adoption, he believes most companies don't need a smarter model. The models are already smart enough. The gap is context they don't absorb - the things an employee who's worked at a company for five years learned by osmosis. If the frontier stopped advancing today, he thinks it wouldn't meaningfully change the value most are extracting from AI anyway. In conversation with a16z's Martin Casado and Sarah Wang: 00:00 Intro 00:48 Why Ali places the AI risk near zero 05:05 The word "pacing" was a mistake 10:50 What 10k agents and $100m can do 12:20 What would change his mind on AI risk 14:20 More GPUs, more ways to fail 18:05 US export controls on PlayStations 20:10 The damage everyone expected by now 24:30 Public vulnerabilities weaponized in hours 30:15 Why labs can't grade each other 37:15 Why most of RSI isn't actually RSI 40:05 Why nobody really needs a smarter model 41:50 The AI use cases nobody argues about 47:30 Google Search solved this 25 years ago 50:55 Nobody has privileged knowledge now 55:10 Same model, new harness, 2x cost 58:15 Open source: 5% of spend, 60% of tokens 1:05:30 90% of new databases are created by agents YouTube: Databricks martin_casado Sarah Wang

a16z

327,508 次观看 • 5 天前

🦙 ollama is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan (Jeffrey Morgan) a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of Lightcone Podcast, Jeff joins Garry Tan, Jared Friedman, Diana, and Harj Taggar to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. 00:43 — The Shift to Open Models 03:03 — How AI Agents Are Driving Token Usage 05:31 — Are Open Models Catching Up? 08:26 — What Happens When a New Model Launches 11:31 — Ollama as an Operating System for AI 14:05 — The New Opportunities Above the Model Layer 18:19 — Why 80–90% of Enterprise Tokens Could Be Open 20:57 — The Future Is Local and Cloud 26:40 — Why AI Is Coming Back to Your Computer 28:56 — The Coming Era of Unlimited Tokens 32:30 — Do We Still Need a “God Model”? 33:41 — Open Models and Geopolitics 36:14 — The Origins of Ollama 40:36 — Two Years Lost in the Wilderness 42:39 — The Pivot That Changed Everything 47:02 — How Ollama Found a Business Model 49:43 — Why Second-Time Founders Did YC

Y Combinator

312,329 次观看 • 19 天前