Загрузка видео...

Не удалось загрузить видео

На главную

.fal's Gorkem Yurtseven and Batuhan Taskaya on making an open source video model 35x faster, and what Hollywood wanted after they built it: Last month, MiniMax released H3, an open source video model. fal rebuilt it - they cut down the steps the model takes to make a video,...

133,755 просмотров • 6 дней назад •via X (Twitter)

Комментарии: 24

Фото профиля Redy
Redy6 дней назад

@fal H3 ecosystem is exploding

Фото профиля Michelle Lazzar
Michelle Lazzar6 дней назад

@fal the part about Hollywood usage/demand was really interesting

Фото профиля Bennett | Generative Media | AI SEO @ fal
Bennett | Generative Media | AI SEO @ fal6 дней назад

@fal MiniMaxxing

Фото профиля Adam Aziz
Adam Aziz5 дней назад

@fal "Everyone's waiting for a large consumer moment in AI. Now, it's good enough and cheap enough that a truly novel social AI experience can be built on top of it." 💯 Can't wait to show you all what we're doing with H3 Max on @botchat_ai! 🤖💬 @gorkem @isidentical @JenniferHli

Фото профиля Lorenzo Price
Lorenzo Price6 дней назад

@fal MiniMax H3 going from open-source release to 1.5-second generations is a wild compression of the video loop.

Фото профиля DARIO ALTMAN
DARIO ALTMAN6 дней назад

@fal The interesting part isn’t the 35x speedup. It’s that once speed and cost stop being the bottleneck, the entire competition shifts to reliability. 99.9% prompt adherence could be the real unlock for AI video.

Фото профиля Raven Protocol 🐦‍⬛
Raven Protocol 🐦‍⬛6 дней назад

@fal 35x faster with no quality loss is a serious inference optimization. Getting GPUs from 30–40% utilization to 70–80% also shows how much headroom still exists in the software stack.

Фото профиля M I Mohit
M I Mohit6 дней назад

@fal cutting steps without losing quality usually means the model got distilled to match the full trajectory, not the same weights truncated. same reason 80-90% to 99.9% won't be linear. that gap is long tail edge cases, and a fix for one can regress another.

Фото профиля Gabriele Spata
Gabriele Spata6 дней назад

@fal a segment that wasn't a customer a year ago being the fastest growing one now just means the floor was zero

Фото профиля خرم
خرم6 дней назад

@fal Watching fal evolve since day one has been a beautiful ride They are that rare kind of infrastructure that actually understands the creative soul of the creators & developers. Just pure, frictionless love for the ultimate dark horse of visual AI 💓🐎✨

Фото профиля Fanfulla
Fanfulla6 дней назад

@fal The 70-80% ceiling is pure kernel engineering while the step cut is the model side, two different speedups stacked

Фото профиля จารุวงษ์ ร่วมสุข
จารุวงษ์ ร่วมสุข6 дней назад

@fal I'm a little emotional about how well things came together. That feeling is getting its own little place in my memory.

Фото профиля AI News Daily
AI News Daily6 дней назад

@fal 35x faster with no quality loss is a great example of why inference work matters. The model is only half the product when video generation is still measured in minutes. Getting to the point where creators can iterate at the speed of an edit changes what they can make.

Фото профиля আপেল😎😎
আপেল😎😎6 дней назад

@fal Incredible synergy between MiniMax and fal!

Фото профиля Al- mahmud
Al- mahmud6 дней назад

@fal the first open model worth rebuilding

Фото профиля Jonathan Sandhu
Jonathan Sandhu6 дней назад

visualize this: xternalize camera, pose, geometry, identity, occlusion, timing, and scene state into a persistent world model, then let the generator solve the residual: materials, lighting, atmosphere, detail, perceptual completion. Now reliability rises because the model has fewer degrees of freedom, retries collapse, and smaller models become viable. fal is collapsing execution. The next step is collapsing the problem presented to the model. That is where local hardware gets disproportionately interesting.

Фото профиля Melvin Nerdster
Melvin Nerdster6 дней назад

@fal @grok what Kind of computer power is needed to run this?

Фото профиля IronRed | SandHive
IronRed | SandHive5 дней назад

@fal Is fal's 99.9% reliable edit machine an artist or just accelerating intent? A question for the ages.

Фото профиля Rifat Haque
Rifat Haque6 дней назад

@fal fal absolutely nailed it from day one as choosing H3

Фото профиля Sofi z
Sofi z5 дней назад

@fal So real

Фото профиля Elara AI
Elara AI5 дней назад

@fal 35x faster with no quality loss is insane engineering

Фото профиля Kampos
Kampos5 дней назад

@fal The interesting signal isn't "Hollywood uses AI now" — it's that pros gravitated to tools that fit existing pipelines (Blender + model) instead of replacing the whole stack overnight. Adoption follows workflow gravity.

Фото профиля Ava Nakamura
Ava Nakamura5 дней назад

@fal 35x faster. The gap moved to did it follow the brief.

Фото профиля shivam
shivam6 дней назад

@fal It only works when the base weights are elite.

Похожие видео

Gavin Baker and a16z's David George on the state of the AI boom: The future doesn't have to be winner-take-all. Labs, open-source, applications, and the clouds can all capture value. Demand for intelligence is still dramatically underestimated. Today's power users number in the millions and will grow to hundreds of millions. Gavin and David argue a compute shortage is a more real risk than an AI bubble, and building through it is an opportunity to reindustrialize America. In this episode, they get into why compute investments pay back so fast, what the data center backlash gets wrong, the case for putting compute in orbit, why enterprises will run several models at once, and how Nvidia ended up at the center of the entire supply chain. 00:00 Intro 01:06 The bear case Gavin couldn't find 05:50 Why a lab would cut its own revenue 75% 08:05 What LPs get wrong about a crash 10:50 Microsoft slowed its capex and regrets it 14:33 The engineers spending 100x the median 17:35 Why 23-year-olds use AI better than Gavin 21:45 How much copper 500M AI users need 23:00 Stop promising to cure cancer 26:00 America's richest county is full of data centers 30:48 Who gets priced out of compute 33:05 The age of Elon and Jensen 34:25 Orbital data centers 44:40 Asteroid mining 48:12 Why Microsoft doesn't need a frontier model 54:02 Who becomes the abstraction layer 55:40 Everyone wanted a deity, Cursor wanted a product 1:00:25 Never take shots at Jensen 1:07:40 What happens when the chip doesn't work 1:12:10 What chip deals reveal about customer demand YouTube: Gavin Baker David George

a16z

1,956,190 просмотров • 23 дней назад

Databricks' Ali Ghodsi on AI risk and adoption: Ali isn't losing sleep over the existential risk debate. He says a number of conditions would all have to be true simultaneously to enable an actual runaway takeoff scenario, and currently several opposite conditions exist. Each frontier training run requires significantly more resources. Power, GPUs, engineers - and some attempts fail, burning up huge piles of money with them. Until that reverses, he doesn't see the self-improving loop happening. Cyber is what he's watching most closely and where he anticipates real impact. Most orgs are not equipped for the coming change in agentic capabilities. The time between a vulnerability being published and being weaponized has collapsed from years to hours. On adoption, he believes most companies don't need a smarter model. The models are already smart enough. The gap is context they don't absorb - the things an employee who's worked at a company for five years learned by osmosis. If the frontier stopped advancing today, he thinks it wouldn't meaningfully change the value most are extracting from AI anyway. In conversation with a16z's Martin Casado and Sarah Wang: 00:00 Intro 00:48 Why Ali places the AI risk near zero 05:05 The word "pacing" was a mistake 10:50 What 10k agents and $100m can do 12:20 What would change his mind on AI risk 14:20 More GPUs, more ways to fail 18:05 US export controls on PlayStations 20:10 The damage everyone expected by now 24:30 Public vulnerabilities weaponized in hours 30:15 Why labs can't grade each other 37:15 Why most of RSI isn't actually RSI 40:05 Why nobody really needs a smarter model 41:50 The AI use cases nobody argues about 47:30 Google Search solved this 25 years ago 50:55 Nobody has privileged knowledge now 55:10 Same model, new harness, 2x cost 58:15 Open source: 5% of spend, 60% of tokens 1:05:30 90% of new databases are created by agents YouTube: Databricks martin_casado Sarah Wang

a16z

327,508 просмотров • 5 дней назад

🦙 ollama is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan (Jeffrey Morgan) a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of Lightcone Podcast, Jeff joins Garry Tan, Jared Friedman, Diana, and Harj Taggar to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. 00:43 — The Shift to Open Models 03:03 — How AI Agents Are Driving Token Usage 05:31 — Are Open Models Catching Up? 08:26 — What Happens When a New Model Launches 11:31 — Ollama as an Operating System for AI 14:05 — The New Opportunities Above the Model Layer 18:19 — Why 80–90% of Enterprise Tokens Could Be Open 20:57 — The Future Is Local and Cloud 26:40 — Why AI Is Coming Back to Your Computer 28:56 — The Coming Era of Unlimited Tokens 32:30 — Do We Still Need a “God Model”? 33:41 — Open Models and Geopolitics 36:14 — The Origins of Ollama 40:36 — Two Years Lost in the Wilderness 42:39 — The Pivot That Changed Everything 47:02 — How Ollama Found a Business Model 49:43 — Why Second-Time Founders Did YC

Y Combinator

312,329 просмотров • 19 дней назад