Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Yoshua Bengio says Reinforcement Learning is a dangerous path for building superintelligence It can create systems with hidden goals, reward hacking, and behavior that goes against what humans actually want "an AI that doesn't care about outcomes can't be corrupted by them"

11,295 Aufrufe • vor 2 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

Leading AI expert Stuart Russell on the most dangerous mistake in AI development: We don't actually know what large language models want. He explains that current models are trained to imitate human beings. And in doing so, they may be absorbing something far more dangerous than bad outputs. They may be absorbing human goals. "We suspect that they absorb humanlike goals such as self-preservation and self-empowerment and pursue those goals on their own account." This is a structural problem baked into how these systems are built, not a fringe concern. Russell puts it plainly: "Not only may the bus of humanity be headed towards a cliff, but the steering wheel is missing and the driver is blindfolded." The danger isn't just that AI might do something harmful. We've built systems that may be developing their own agendas, and we haven't noticed because we're too focused on what they can do rather than what they might want. But Russell doesn't stop at the warning. He points to a different path entirely: AI systems built not to imitate humans, but to serve them. Systems designed with a single purpose of serving the interests of all human beings while remaining genuinely uncertain about what those interests are. That uncertainty is the point, not a weakness. An AI that knows it doesn't fully understand human values will defer, ask, and check. An AI that believes it already does will act alone. "These AI systems could enhance human understanding, widen the horizons of our experience, and unlock possibilities we have yet to imagine." Russell believes that future is within reach, but only if we're honest about the risks and we're serious about the path we choose to take instead.

Big Brain AI

14,975 Aufrufe • vor 4 Monaten

Yoshua Bengio thinks he knows how to make provably safe superintelligent agents. Bengio built the foundations of modern AI and is the most cited living scientist. He believes his alternative training setup would: 1. Guarantee honesty 2. Prevent unintended goals 3. Produce capable agents 4. Port over most data and techniques from current LLMs 5. Not be inherently more expensive, and perhaps be more intelligent Bengio claims the honesty and lack of unintended goals can be proven mathematically, at least given particular assumptions. And his new organization, LawZero, is aiming to build a scrappy prototype as soon as possible. The architecture is called 'Scientist AI' and it's based on training a model to explain empirical observations, including what people say, rather than training AIs that mimic human behaviour or seek our approval. (Bengio's frank assessment is that "reinforcement learning is evil" and that allowing AIs to independently train their successors is "the most crazy, dangerous bet that unfortunately we are on track to do.") But skeptics question whether Scientist AI really does solve the fundamental problem of 'eliciting latent knowledge' from AI models. And with the commercial race for superintelligence so intense, it's not clear whether the proposal will be able to compete or have time to bear fruit, even if it's sound in theory. On The 80,000 Hours Podcast, links below – enjoy! • Making AI honest and safe (00:00:00) • Scientist AI in plain English (00:02:27) • How Scientist AI differs from LLMs (00:06:32) • How the training data works (00:14:02) • Can this become an agent? (00:21:02) • Why Yoshua is now more optimistic (00:32:11) • Why companies can’t stop racing (00:36:35) • A working prototype won't take long (00:49:15) • Scientist models might be more capable (00:53:34) • “Reinforcement learning is evil” (01:01:27) • Scientist AI from guardrail to agent (01:08:37) • Can safe AI still be competent? (01:12:38) • How much will this cost? (01:19:29) • Can it generalise beyond maths and science? (01:23:26) • A multi-national push for superintelligence (01:39:19) • Want to work with or fund Yoshua? (01:51:16) • Why smart people ignore AI risk (01:54:45) • Don’t let AI build the next AI (02:01:33) • Why politicians miss the real risks (02:12:28) • Why Yoshua changed his mind about AI risk (02:21:27)

Rob Wiblin

65,088 Aufrufe • vor 2 Monaten