Загрузка видео...

Не удалось загрузить видео

На главную

Deep learning works extraordinarily well. And we still largely don't know why. A new paper from Jamie Simon, Daniel Kunin, and 12 co-authors argues that a scientific theory of deep learning is emerging, and coins a name for the emerging field: learning mechanics. We sat down with Jamie and...

18,081 просмотров • 2 месяцев назад •via X (Twitter)

Комментарии: 0

Нет доступных комментариев

Здесь появятся комментарии из оригинального поста

Похожие видео

François Chollet (François Chollet) has spent years asking a different question than most of the AI world. Instead of scaling what already works, he’s trying to understand what intelligence actually is and how to build it from first principles. In this episode of the Lightcone Podcast, he traces that path from his early work on deep learning to the creation of the ARC Prize, and the launch of ARC V3, a new benchmark designed to measure something deeper than performance: the ability to learn, adapt, and reason efficiently in entirely new environments. He explains why today’s systems may be hitting limits, what recent breakthroughs really mean, and why reaching true general intelligence may require a fundamentally different approach. 00:00 - AGI by 2030? 00:31 - Introducing Ndea: A New Path Beyond Deep Learning 01:08 - A New ML Paradigm 01:30 - Replacing neural nets with compact symbolic programs 03:04 - Why Ndea Isn’t Competing With Coding Agents 05:20 - Why Everyone Might Be Wrong About Scaling LLMs 07:22 - Why Coding Agents Suddenly Work So Well 08:50 - The Limits of LLMs in Non-Verifiable Domains 10:48 - What AGI Actually Means (And Why Most Definitions Are Wrong) 13:30 - Why Deep Learning Hits a Wall 14:00 - ARC’s Origin Story 18:20 - ARC Benchmarks Explained: From V1 to V3 22:49 - The RL Loop Powering Coding Agents Today 27:03 - ARC-AGI V3: Measuring “Agentic Intelligence” 31:14 - Inside the ARC Game Studio 35:31 - Could AGI Fit in 10,000 Lines of Code? 44:01 - Building Ndea: From Idea to Compounding Research Stack 46:46 - The Future of ARC: Benchmarks That Evolve With AI 47:21 - Why There’s Still Huge Opportunity for New AI Paradigms 53:37 - How to Build a Breakout Open Source Project - Lessons From Keras 56:39 - Advice For How To Think About AI

Y Combinator

151,332 просмотров • 3 месяцев назад

I had a fantastic time discussing with the learning legend Justin Skycak from Math Academy about learning math in the modern age. we've talked about his quite impressive self-learning journey (3000h of math in high school) all the way to how he hand curated the initial knowledge graph for math academy to make that process more efficient. great lively 3h discussion here are the chapters: 0:00:00 - intro: 0:02:10 - justin background 0:05:45 - 3000h math self study in high school 0:11:45 - what a day looked like for that 3000h stretch 0:16:10 - meta-learning vs pure math learning 0:21:50 - when did you get into cognitive neuro? 0:29:55 - how did the fundamental math helped in your research projects 0:43:10 - what does the math academy learning system looks like 0:47:34 - how did you guys build the 2000 topic knowledge graph 1:01:15 - would LLM be useful as an interface to that knowledge graph for the students? 1:10:46 - how does the FIRe spaced repetition algorithm works? 1:17:34 - does the same knowledge graph structure would work for physics? or other topic?: 1:34:05 - how do you understand the subject vs the curiculum 1:35:50 - is there a connection between studying math and learning a sport? 1:42:00 - do you think in math doing and teaching requires different skills? 1:56:25 - could you get understanding without automaticy? 2:05:35 - do you see any upside of confusion in learning? 2:14:11 - learning math as an adult? 2:19:20 - how to fill the motivation gap after learning the fundamental? 2:24:10 - how should teaching math for kids and adults balance fundamentals and creativity? 2:33:55 - is it ever too late to learn math seriously? 2:46:00 - mastery learning vs ultra learning 2:51:30 - top-down vs bottom-up 2:53:40 - mastery learning for domain without a structured hierarchical structure? 2:56:30 - neurodivergence / adhd for structured math learning? 3:06:20 - amateur mathematician augmented with technology will be able to contribute to research? 3:14:37 - what are you most excited about right now in term of learning enjoy!

Yacine Mahdid

57,320 просмотров • 3 месяцев назад

New episode with Dr. Konrad Kording (Kording Lab 🦖), professor of bioengineering and neuroscience at the University of Pennsylvania (Penn) and co-director of CIFAR's Learning in Machines & Brains program (CIFAR). Konrad works at the intersection of causality, machine learning, and neuroscience, building rigorous methods for causal reasoning when experiments aren't possible — and challenging how researchers interpret neural data and build AI. Konrad argues the most promising path to understanding how the brain works is to read the brain’s wiring directly, down to the molecular detail of each connection, and to build compilers and simulations to understand the brain’s computation directly. In this episode we go deep into how neurons work, how neurons wire together, and how organic and artificial neural networks differ. We discuss why organic neurons are doing much more; how a model of a single organic neuron can solve MNIST — computing more like a 3-layer artificial neural network; how the brain might learn by solving credit assignment with only local signals; how to approximate backprop without a global algorithm; why AI and humans are intelligent along different dimensions; why Konrad isn’t very worried about AI replacing us; economic models of intelligence and physical work; and much more. Konrad is a brilliant, contrarian thinker who explains complex concepts very intuitively. It is a solid computational neuroscience primer. I hope you enjoy this conversation as much as I did! Other links to this episode and references below. Chapters 00:00:00 Introduction 00:01:01 How organic neurons work 00:24:13 How the brain learns: circuits and credit assignment 00:45:29 Recording the brain 00:52:47 Why simulating brains is hard 01:05:00 A new approach: connectomes and compilers 01:21:00 Why simulate brains? 01:29:50 How AI and human intelligence differ 01:41:04 Evolution, intelligence and AI risk 01:52:42 Robotics, causality, and the roots of intelligence 02:05:53 AI for science and scientific rigor 02:13:05 The economics of intelligence 02:27:50 A hopeful future

Juan Benet

49,297 просмотров • 25 дней назад

E133: Sam Blackshear - How Libra Sparked the Move Language and Why Sui Is the Real Endgame! Sam Blackshear is the Co-founder and CTO of MystenLabs.sui , the company behind the Sui, and the Creator of the Move programming language that's revolutionizing smart contract development. Timestamps: 0:00 Introduction 1:54 Partnerships: Jupiter, KAST (old), , Sui, Mantle, Forza! BTC 2:44 The Power of Preparation 5:14 Discipline Behind the Podcast 6:27 Translating Thought Into Code 8:19 Who is Sam Blackshear? 9:27 Choosing What Truly Matters 10:25 Self-Custody with Trezor 11:18 Crypto vs. AI Thinking 12:28 The Power of Support 16:04 From Court Dreams to Reality 17:55 Challenging the Limits of Code 22:52 Chose Learning Over a Job 24:09 The Internship That Changed Everything 27:35 PhD Skills Meet Facebook 29:05 Entering Crypto Through Facebook 32:15 Why Libra Needed Move 33:42 Solving Scarcity in Code 36:37 Bitcoin & Ethereum Mistakes 38:45 Creating a New Language 41:41 Problem-Driven Innovation 44:47 Avoiding Analysis Paralysis 48:11 What is Unstructured Thinking? 50:50 Why Unstructured Thinking Works 53:14 Future of Crypto Protocols 54:21 Why Move is the Best Programming Language 55:00 What is the Sui Network? 56:25 What Makes Sui Different? 57:05 Why is Sui The Best Blockchain? 59:05 Managing Energy Long-Term 1:01:02 Satisfaction Without Closure 1:03:37 90% Love, 10% Grind 1:05:37 Mental State of Surfing 1:06:47 Non-Consensus Beliefs 1:07:34 What is Memory Safety? 1:09:07 What Was the Equifax Hack? 1:10:53 Rethinking Software Safety 1:12:00 Right Dose of Regulation 1:13:00 Biggest Prediction for the Next 24 Months? 1:14:02 Scaling Crypto Developers 1:16:07 Concluding Remarks

MR SHIFT 🦁

164,259 просмотров • 11 месяцев назад

Today’s guest on the Free Radicals podcast is José Luis Ricón Fernández de la Puente, head of theory Retro Biosciences and blogger Jose is a prolific blogger, covering a wide breadth of topics across economics, philosophy, progress studies, science funding, much more, and of course longevity. His works have been published by a16z, Works in Progress and Adam Smith Institute. Our conversation is wide ranging, spanning a deep dive on Retro’s work to replace and engineer microglia to rejuvenate the brain and how our cells have the ability to turn back the aging clock but choose not to. We also covered the technological stagnation and why biological engineering is the new frontier of progress, as well as philosophical topics like transhumanism and how a future of total biological control might impact our values and way of life. Be sure to follow me and Eric Dai to stay up to date on the latest news in longevity biotech! And Special thank you to NFX & omri_drory for lending us their beautiful podcasting studio! 0:00 Intro 2:49 What is aging & why cells have a tough choice to make 9:02 When cells choose to reverse aging themselves 12:43 Cellular vs Organismal Aging & the magic wand experiment 18:31 What is reprogramming 22:40 How reprogramming plays a role in DNA damage repair 25:42 Do we already know how to cure aging? FOXO3! 28:37 How to cut through the complexity of interconnected biology 32:32 Why transcription factors are so great for intervening 36:49 Does a rejuvenation program exist already in the genome 38:51 Michael Levin: from thinking in terms of genes to morphogenesis 48:14 Tech stagnation and why physics is cooked 55:03 Why doesn't the world look more futuristic 57:26 Transhumanism & asking ourselves what we want out of life 1:05:34 Do we need war for technological progress 1:09:31 Government role in science funding 1:15:06 How Jose became the Head of Theory at Retro 1:24:53 How AI might put software engineers out of a job, and push them towards biotech 1:27:28 What it takes to get a flywheel in biotech 1:29:04 Rejuvenation vs Prevention 1:34:23 Aging is the coolest hardest problem to work on 1:36:09 What does it take to cure aging 1:42:25 Delivery mechanisms for genetic therapies 1:48:53 Retro's work to replace microglia and engineer them outside the body 1:59:10 Consciousness 2:00:39 Jose's Origin Story

Daniel Shur

13,616 просмотров • 4 месяцев назад

if you're struggling on where to start learning ML, here’s a playlist of 30 youtube videos to learn machine learning fundamentals from scratch "Machine Learning: Teach by Doing" is a solid choice to learn both theory and code. (1) Introduction to Machine Learning Teach by Doing: (2) What is Machine Learning? History of Machine Learning: (3) Types of ML Models: (4) 6 steps of any ML project: (5) Install Python and VSCode and run your first code: (6) Linear Classifiers Part 1: (7) Linear Classifiers Part 2: (8) Jupyter Notebook, Numpy and Scikit-Learn: (9) Running the Random Linear Classifier Algorithm in Python: (10) The oldest ML model - Perceptron: (11) Coding the Perceptron: (12) Perceptron Convergence Theorem: (13) Magic of features in Machine Learning: (14) One hot encoding: (15) Logistic Regression Part 1: (16) Cross Entropy Loss: (17) How gradient descent works: (18) Logistic Regression from scratch in Python: (19) Introduction to Regularization: (20) Implementing Regularization in Python: (21) Linear Regression Introduction: (22) Ordinary Least Squares step by step implementation: (23) Ridge regression fundamentals and intuition: (24) Regression recap for interviews: (25) Neural network architecture in 30 minutes: (26) Backpropagation intuition: (27) Neural network activation functions: (28) Momentum in gradient descent: (29) Hands on neural network training in Python: (30) Introduction to Convolutional Neural Networks (CNNs):

ℏεsam

108,861 просмотров • 1 год назад

a playlist of 30 youtube videos to learn machine learning fundamentals from scratch if you're struggling on where to start learning ML, this list goes this "Machine Learning: Teach by Doing" is a solid choice to learn both theory and code. (1) Introduction to Machine Learning Teach by Doing: (2) What is Machine Learning? History of Machine Learning: (3) Types of ML Models: (4) 6 steps of any ML project: (5) Install Python and VSCode and run your first code: (6) Linear Classifiers Part 1: (7) Linear Classifiers Part 2: (8) Jupyter Notebook, Numpy and Scikit-Learn: (9) Running the Random Linear Classifier Algorithm in Python: (10) The oldest ML model - Perceptron: (11) Coding the Perceptron: (12) Perceptron Convergence Theorem: (13) Magic of features in Machine Learning: (14) One hot encoding: (15) Logistic Regression Part 1: (16) Cross Entropy Loss: (17) How gradient descent works: (18) Logistic Regression from scratch in Python: (19) Introduction to Regularization: (20) Implementing Regularization in Python: (21) Linear Regression Introduction: (22) Ordinary Least Squares step by step implementation: (23) Ridge regression fundamentals and intuition: (24) Regression recap for interviews: (25) Neural network architecture in 30 minutes: (26) Backpropagation intuition: (27) Neural network activation functions: (28) Momentum in gradient descent: (29) Hands on neural network training in Python: (30) Introduction to Convolutional Neural Networks (CNNs):

ℏεsam

117,570 просмотров • 1 год назад

The most capable individuals are learning and shipping 1000x faster using AI. I sat down with Gabriel Petersson () - a high school dropout who self-taught Math & ML with AI and is now an AI Researcher at OpenAI. We deep dive into the most important traits to thrive in the age of infinite leverage: – How to learn hard skills insanely fast with AI – Dropping out in Sweden → America on the O-1 Extraordinary visa - How to get hired by the best when you are a nobody – How to turn prompting into a daily habit – Why America is still the launchpad for ambitious builders - Breaking down Gabriel's most iconic tweets Curiosity + Agency + AI will take you further than any legacy credential Timestamps: 01:12 - Swedish high-school dropout → OpenAI research scientist 03:04 - First startup & sales hacks: door-knocking, instant A/B tests 05:28 - Couch-life grind & learning to code under pressure 08:16 - Learning by doing: top-down vs bottom-up; what schools miss about AI 12:10 - Recursive learning with ChatGPT: ELI5, intuition, gap-filling 16:20 - Prompting habits, make AI your always-on tutor 22:16 - Research workflow: papers, code inserts; “shortcuts to foundations” 27:13 - Feedback obsession, elite code reviews, AI as reviewer 33:12 - Path to the US🇺🇸 Midjourney, getting the O-1 Extraordinary visa, shipping demos that signal skill 39:29 - bypass recruiters via proof of work, risk-free trials & academia hot takes 55:06 - why San Francisco; visas & how to move; final pep talk

Sigil Wen

1,297,897 просмотров • 7 месяцев назад

Inside Nemotron and NVIDIA's AI lab: my conversation with Bryan Catanzaro (Bryan Catanzaro). NVIDIA is a chip company. So why does it put hundreds of researchers on building AI models - and then give them away for free? We go deep into the Nemotron models, what it takes to build a top AI lab, and the future of frontier AI. 01:33 - Is open source AI catching the frontier? 05:29 - Do closed labs blocking distillation slow open source down? 07:42 - Is the US falling behind China? 10:30 - Why companies actually choose open models 12:39 - A "crazy" 2008 bet: machine learning on GPUs 15:33 - Working with Andrew Ng and Dario Amodei at Baidu 17:41 - Coming back to NVIDIA: DLSS and the birth of Megatron 21:55 - The real reason NVIDIA builds its own models 24:28 - Is Moore's Law really dead? 33:37 - The Nemotron family: Nano, Super, Ultra 35:09 - Built for agents: why NVIDIA bets on speed 36:02 - How you train a 550B model in 4 bits 39:25 - Hybrid Mamba-Transformer, explained simply 42:31 - Mixture of experts, and why NVIDIA built NVL72 around it 47:26 - Why a 1-million-token context window matters 49:26 - Multi-token prediction: how the model predicts 5 tokens at once 52:47 - Multi-teacher distillation: teaching one model from many 58:01 - Where reinforcement learning goes next 01:00:16 - Inside NVIDIA's research org: "the mission is the boss" 01:04:03 - How NVIDIA decides who gets the GPUs 01:10:53 - Why NVIDIA still feels entrepreneurial after 33 years 01:12:58 - Why Bryan doesn't believe in the singularity 01:17:50 - The AI backlash 01:19:18 - The controversial case: open AI is safer than closed

Matt Turck

56,381 просмотров • 18 дней назад