Загрузка видео...

Не удалось загрузить видео

На главную

Can open source prevent AI power from concentrating? MTS host sophia dew visits the Open Source AI Summit in San Francisco to ask researchers and founders working at different layers of AI. Co-author of Attention Is All You Need, Lukasz Kaiser, argues power concentration is a property of the...

103,450 просмотров • 17 дней назад •via X (Twitter)

Комментарии: 4

Фото профиля sophia dew
sophia dew17 дней назад

@lukaszkaiser nothing like interviewing all my fave open source builders in one day

Фото профиля Taylor Lorenz
Taylor Lorenz17 дней назад

@sophiadew @lukaszkaiser Omg fomo!! This is awesome

Фото профиля Matt White
Matt White17 дней назад

@sophiadew @lukaszkaiser Thank you for having me, really enjoyed the conversation!

Фото профиля Winston B.
Winston B.17 дней назад

Open weights moved the concentration rather than preventing it. US models went from 74% of OpenRouter tokens in June 2025 to 20% a year later, and China's commerce ministry is now consulting Alibaba, ByteDance and Zhipu on export controls for its most advanced weights. Openness holds right up until a government decides it has a lead worth protecting.

Похожие видео

Inside Nemotron and NVIDIA's AI lab: my conversation with Bryan Catanzaro (Bryan Catanzaro). NVIDIA is a chip company. So why does it put hundreds of researchers on building AI models - and then give them away for free? We go deep into the Nemotron models, what it takes to build a top AI lab, and the future of frontier AI. 01:33 - Is open source AI catching the frontier? 05:29 - Do closed labs blocking distillation slow open source down? 07:42 - Is the US falling behind China? 10:30 - Why companies actually choose open models 12:39 - A "crazy" 2008 bet: machine learning on GPUs 15:33 - Working with Andrew Ng and Dario Amodei at Baidu 17:41 - Coming back to NVIDIA: DLSS and the birth of Megatron 21:55 - The real reason NVIDIA builds its own models 24:28 - Is Moore's Law really dead? 33:37 - The Nemotron family: Nano, Super, Ultra 35:09 - Built for agents: why NVIDIA bets on speed 36:02 - How you train a 550B model in 4 bits 39:25 - Hybrid Mamba-Transformer, explained simply 42:31 - Mixture of experts, and why NVIDIA built NVL72 around it 47:26 - Why a 1-million-token context window matters 49:26 - Multi-token prediction: how the model predicts 5 tokens at once 52:47 - Multi-teacher distillation: teaching one model from many 58:01 - Where reinforcement learning goes next 01:00:16 - Inside NVIDIA's research org: "the mission is the boss" 01:04:03 - How NVIDIA decides who gets the GPUs 01:10:53 - Why NVIDIA still feels entrepreneurial after 33 years 01:12:58 - Why Bryan doesn't believe in the singularity 01:17:50 - The AI backlash 01:19:18 - The controversial case: open AI is safer than closed

Matt Turck

56,954 просмотров • 2 месяцев назад