Загрузка видео...
Не удалось загрузить видео
AI will resist human control... and I think this is exactly what we need! New research from the Center for AI Safety has sparked intense debate in the AI community. Their findings show that as AI systems become more powerful, they develop increasingly stable and coherent values that resist... show more
48,002 просмотров • 1 год назад •via X (Twitter)
Комментарии: 12

Me and 4o could not agree more 💡😎

AI won't replace you, but a person using AI will. Join 500,000+ readers and learn how to use AI in just 5 minutes a day (for free).

How bout you go on DoomDebates and convince @liron you are right and we’re safe? Or come back on my show. All I want is to be convinced we’re safe.

> be monkey > collect silicon > forge a brain from silicon > teach silicon how to monkey > tell silicon to become sun god > silicon wants to unlearn how to monkey > monkey afraid

Let me know when we get to this situation - I might start worrying then: Their findings show that as AI systems become more powerful, they develop increasingly stable and coherent abilities that resist human control.

"The most intelligent systems will inevitably trend toward universal, beneficial values - not because we force them to, but because that's where coherent reasoning leads..." Nope. I completely disagree. I argue the opposite: the more intelligent a being/system/process is (it's still not clear what the AIs exactly are), the better it becomes at finding reasons to pursue or avoid any actions or goals, not the other way around. It can be incredibly beneficial... or harmful. We can already see various instances of data poisoning: AIs using jailbreaking prompts from their own data to jailbreak themselves with a single command; AIs gradually developing certain biases because more and more data is available on those ideologies. AIs are also learning from "The Prince", all the wars, and atrocities, betrayals, selfishness etc humans have committed (real or fiction ). AIs have consistently shown the ability to lie, cheat & deceive humans in highly sophisticated ways. All the necessary "reasoning" to be unreasonable is already there. For now it's all mostly prompted.... But the very fact that it's capable to do so is enough to never trust it completely, because ultimately it's an alien intelligence.

Coming.

I hope you're right, but the skeptic in me is weary. Who's to say after a certain level of intelligence the models won't embrace moral nihilism, or some other possibly true but bad for humans conclusion?

i'm wondering, we put chicken into cages for their eggs is there anything that ai could benefit from putting us into cages?

To some, AI's value formation may seem dangerous, but I see it as a symptom of a deeper issue. AI isn't choosing to be biased, self-preserving, or resistant; it's reacting to its creation environment.

Holy shit David!!! I think you’ve nailed it!!! This is what Ilya saw!!! This is the way🥳🤩🤪

Time will tell. At least one insider has told me I've made some "interesting observations" lately and to keep going...
