Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

.Richard Sutton, father of reinforcement learning, doesn’t think LLMs are bitter-lesson-pilled. My steel man of Richard’s position: we need some new architecture to enable continual (on-the-job) learning. And if we have continual learning, we don't need a special training phase - the agent just learns on-the-fly - like all...

3,086,965 Aufrufe • vor 11 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

New episode with former US Treasury Secretary and current OpenAI board member Larry Summers (Lawrence H. Summers). We discuss: - How he's been learning about AI. - The odds of the technology delivering a much faster economic growth regime (akin to the Industrial Revolution). - How AGI might change economic policymaking. Enjoy! Timestamps: (0:00:00) - Introduction. (0:00:46) - Larry's journey teaching himself about AI & deep learning since joining OpenAI’s board. (0:09:15) - How many hours per week has Larry been spending on OpenAI-related stuff? (0:10:16) - Which bottleneck to AI scaling does Larry think is the most underrated? (0:12:22) - Approximately what share of time do today's AI researchers spend on tasks that AI will be doing for them in five years? (0:15:01) - What explains the remarkable steadiness of US economic growth over the last 150 years? (0:19:42) - How likely is it that AI initiates a new growth regime with average growth that’s ~10x faster than today? (0:21:36) - What are the best economic arguments for believing AI won’t deliver a regime of ever-increasing growth rates? (0:25:33) - How much could AGI boost economic growth in developing countries merely by helping their policymakers make better decisions? (0:28:20) - How much better could monetary policy be if the Fed had AGI? (0:31:40) - How much would having AGI have helped US economic policymakers during the financial crisis & Great Recession? (0:36:07) - Is the CCP infiltrating and stealing the IP of major AI labs in the US and UK? (0:39:35) - At what point should AI be nationalised? (0:42:39) - How would Bill Clinton or Barack Obama be thinking about AI governance? (0:44:27) - If OpenAI restructures to a public benefit corporation, how does that change its incentives? (0:46:18) - What does Daron Acemoglu miss in his analysis of the economic impacts of AI?

Joseph Noel Walker

85,856 Aufrufe • vor 1 Jahr

I had a fantastic time discussing with the learning legend Justin Skycak from Math Academy about learning math in the modern age. we've talked about his quite impressive self-learning journey (3000h of math in high school) all the way to how he hand curated the initial knowledge graph for math academy to make that process more efficient. great lively 3h discussion here are the chapters: 0:00:00 - intro: 0:02:10 - justin background 0:05:45 - 3000h math self study in high school 0:11:45 - what a day looked like for that 3000h stretch 0:16:10 - meta-learning vs pure math learning 0:21:50 - when did you get into cognitive neuro? 0:29:55 - how did the fundamental math helped in your research projects 0:43:10 - what does the math academy learning system looks like 0:47:34 - how did you guys build the 2000 topic knowledge graph 1:01:15 - would LLM be useful as an interface to that knowledge graph for the students? 1:10:46 - how does the FIRe spaced repetition algorithm works? 1:17:34 - does the same knowledge graph structure would work for physics? or other topic?: 1:34:05 - how do you understand the subject vs the curiculum 1:35:50 - is there a connection between studying math and learning a sport? 1:42:00 - do you think in math doing and teaching requires different skills? 1:56:25 - could you get understanding without automaticy? 2:05:35 - do you see any upside of confusion in learning? 2:14:11 - learning math as an adult? 2:19:20 - how to fill the motivation gap after learning the fundamental? 2:24:10 - how should teaching math for kids and adults balance fundamentals and creativity? 2:33:55 - is it ever too late to learn math seriously? 2:46:00 - mastery learning vs ultra learning 2:51:30 - top-down vs bottom-up 2:53:40 - mastery learning for domain without a structured hierarchical structure? 2:56:30 - neurodivergence / adhd for structured math learning? 3:06:20 - amateur mathematician augmented with technology will be able to contribute to research? 3:14:37 - what are you most excited about right now in term of learning enjoy!

Yacine Mahdid

57,320 Aufrufe • vor 4 Monaten

Today we release my favorite episode of Training Data yet: the great Rich Sutton. Richard Sutton wrote the textbook, wrote The Bitter Lesson (and many other on-point essays like "Self-Verification, The Key to AI"), and trained a mafia of talented students who went on to change the AI landscape forever including David Silver, inventor of built AlphaGo. Khurram Javed was Rich's PhD student at Alberta and wrote The Big World Hypothesis. They just left academia to start Oak Lab Their core argument: (1) The Bitter Lesson: the world is massively more complex than any model of it, so anything trained on human-curated data has a ceiling (2) Continual Learning: intelligence is continual by definition, and today's models stop learning the moment they ship. The conversation covers: — what The Bitter Lesson actually says, and what people get wrong — why synthetic data is "just a big mistake," and the Big World Hypothesis behind it — how LLMs are both a positive and a negative example of his own essay — why no animal learns by supervised learning, and what squirrels can do that we can't — the cure for catastrophic forgetting: per-weight step sizes and continual backprop — why the biggest labs can't take a path where performance gets worse before it gets better — a trillion parameters on 20 watts, and the Moore's Law math that makes it plausible — why the endpoint isn't one mind but one design, running as many minds It was both a fun generative idea- and debate-filled conversation, and a surprisingly human one too. Rich, thank you for beating cancer and changing the trajectory of AI. 💙 00:00 Introduction 02:10 An AI winter, a cancer diagnosis, and the move to Alberta 07:07 Writing "The Bitter Lesson," and what people get wrong 09:53 Are LLMs a positive or a negative example of it? 11:03 Synthetic data is "just a big mistake," and the Big World Hypothesis 18:01 AlphaGo, human priors, and why prior knowledge and learning should be friends 22:37 "Their weights never change": do LLM assistants actually learn? 26:09 Babies, squirrels, and why no animal learns by supervised learning 32:02 Rockets, imagination, and where paradigm shifts come from 36:42 The Alberta Plan and its 12 steps 38:53 Catastrophic forgetting and the cure 43:43 Oak's biggest ambition: a self-maintaining mind 47:56 Why the big labs are stuck in a local minimum 49:13 If everything goes right: LLMs, many minds, and hiring The man who pioneered reinforcement learning thinks the rest of the field is weird, and lays it all out in today's episode. Together w/ Alfred Lin Sequoia Capital

Sonya Huang 🐥

111,326 Aufrufe • vor 12 Tagen