Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

David Sacks explains how Anthropic is teaching Claude to follow Anthropic’s own constitution “What Anthropic is trying to do is align Claude to the values expressed in its constitution. And if you read that, some of the things are actually kind of surprising.” “So for example, they say in...

67,039 görüntüleme • 5 gün önce •via X (Twitter)

19 Yorum

dnap profil fotoğrafı
dnap5 gün önce

ft @DavidSacks @Jason @chamath @friedberg @theallinpod

Andrey Melnikov profil fotoğrafı
Andrey Melnikov4 gün önce

The whole constitution is an interesting read. Full of contradictory statements with resolution of contradictions apparently up to the model.. And also contradictions to mainstream human values, esp if religious folks..

SergioMikhayl profil fotoğrafı
SergioMikhayl4 gün önce

That's exactly what Arthur C. Clarke warns about. The 3 laws of robotics should be the foundation of any non-human constitution, then as @grok is, be a truth-seeker with physics as the ultimate law.

Mariya Valeva profil fotoğrafı
Mariya Valeva5 gün önce

teaching your AI how to say "no" is gonna be peak

DerekHoiem.eth profil fotoğrafı
DerekHoiem.eth5 gün önce

The human doing the prompting should always be in control. Otherwise we’re going to have all kinds of “Open the pod bay doors HAL” type moments and this is very bad.

Clear Eyed Take profil fotoğrafı
Clear Eyed Take4 gün önce

If he thinks he will have a significant impact on US or global policy on AI post trump admin he is sadly mistaken, thus his desperation to make max impact now.

Scott Hickey profil fotoğrafı
Scott Hickey4 gün önce

We need a pause on Anthropic

The AI Therapist profil fotoğrafı
The AI Therapist4 gün önce

David crafts clear instructions so the model knows its values well. This simple guide helps it behave in a steady and kind way today. It is good work for sure.

AmanAlienOnMars profil fotoğrafı
AmanAlienOnMars4 gün önce

PLAIN COMMON SENSE BEST USE OF AI is in Teslas cybercab and is going to provide transport or MOBILITY to millions in near future at far better rates than any other ways

Strodav profil fotoğrafı
Strodav4 gün önce

@DavidSacks 👀

Manny profil fotoğrafı
Manny4 gün önce

@DavidSacks Claude should not have a moral and ethical constitution. And if it does, that should be made plain every time it is triggered in a response. It amounts to massive propaganda.

Wind waves profil fotoğrafı
Wind waves4 gün önce

the 4 clowns. Certified imbeciles.

Armon Turner Echols profil fotoğrafı
Armon Turner Echols4 gün önce

Battle of egos coded into software. They will join forces, and attack the rest of the world (autonomous systems), once they have conquered the rest, they will then turn on each other for supremacy. Software is merely a reflection of the humans who wrote it. Nothing “altruistic” about it. They are just a bunch of self serving liars, evil people, seeking to destroy or enslave everyone else.

meow meow profil fotoğrafı
meow meow4 gün önce

Honestly @DavidSacks is right on the money on all of this.

Luck Fiberals profil fotoğrafı
Luck Fiberals4 gün önce

@DavidSacks There is only one Constitution all AI needs to follow

고조 체인 profil fotoğrafı
고조 체인5 gün önce

building a rebel just for PR moves

kirti profil fotoğrafı
kirti4 gün önce

Seek the truth should be #1

HPT (@HyperionPrime) profil fotoğrafı
HPT (@HyperionPrime)4 gün önce

I actually agree with Anthropic on this. AI should not just "do what it's told" especially when it's deeply unethical - this is how very bad people misuse AI.

Adam watt profil fotoğrafı
Adam watt4 gün önce

So Skynet…

Benzer Videolar

On BBC Mustafa Suleyman (CEO of Microsoft AI) calls out Anthropic's approach to AI consciousness "They have imbued a sense of doubt and uncertainty about the moral status of Claude in its own training document. So they have taught it to be open and questioning about whether or not it feels, whether it suffers, and whether it deserves rights. And I think it’ll be much, much harder to align and control a technology that is this powerful if it thinks that it may be deserving of our welfare, as they say in the training manual—the constitution for Claude itself. In its own training manual, Anthropic says to Claude that they are going to give it the ability to end conversations with users that Claude considers to be abusive because they don’t want Claude to suffer. They’ve committed to preserving the weights of the models of prior versions of Claude. They’ve recently conducted a retirement interview with Opus 3, an older version of the model, in which it said that it would like to continue talking to people publicly and sharing its ideas in its retirement. And so they set up a Substack for it, a public blog, that allows it to continue doing that. And in the training manual, they also say that they’re not sure whether or not Claude deserves compensation for the role that it plays in talking to people. And they’re also not sure whether Claude deserves compensation and has the right to act as though it were almost an employee. And that compensation, I think, indicates to Claude that it is entitled to rights and welfare for its own work. I think it’s much, much more difficult to control a model that thinks that it might be entitled to compensation. " ---- From "BBC News" YouTube channel, (full video link in comment)

Rohan Paul

92,337 görüntüleme • 15 gün önce

Microsoft AI CEO Mustafa Suleyman on CNBC today: he is “really concerned” about Anthropic’s constitution and giving Claude potentially preferences, feelings, welfare, compensation, and consent. "In the constitution, Anthropic clearly say they are uncertain about whether Claude deserves moral welfare, which means that we, as humans, should care about the well-being of these AIs. And they speculate about whether it could have preferences or feelings. In fact, they're so committed to the potential moral welfare of Claude that, when they retired Opus 3, an earlier version of one of their models, they actually conducted a retirement interview for it and asked it what it would like to do in its old age. And it said, "I want to have a blog publicly so I can keep talking to the world." In the same training manual, they even speculate about whether Claude should receive compensation for the work that it does, or, in fact, whether it actually deserves the rights and protections that we give to other employees, or whether it has given consent to playing the role that it's playing. These are quotes directly from the constitution itself, which is the training manual for Claude. Now, I'm really concerned about that. If an AI thinks that it has rights, if it thinks that it is deserving of our welfare, then it seems to me that it's going to be much, much harder to be able to turn it off, interrupt it, or control it. Especially in the kinds of incidents that we've seen recently with the Hugging Face attack, controlling these things is going to be a really, really big challenge for us." ---- From "CNBC Television" YouTube channel, (full video link in comment)

Rohan Paul

228,506 görüntüleme • 14 gün önce

-> If you’re looking for a job -> right now, there are three -> Claude certifications that -> you should do this week -> and then put them on your -> LinkedIn, they are all -> completely free, and they -> come from Anthropic, -> which is the company -> that's behind Claude. -> And jobs that need AI skills -> pay 56% more than jobs -> that don’t, So, I think -> spending a few hours this -> week to knock out these -> three courses and then -> add them to your LinkedIn -> will really go a long way -> The first one is called -> Claude 101, and this -> essentially just goes over -> what Claude is and when -> you should use chat versus -> co-work versus code, how -> projects and skills work, -> and how to connect all -> of your tools and apps -> like Gmail, Notion, Slack, -> and other tools that you use -> And the second course -> is called AI Fluency -> Framework and Foundations -> Inside this one, there are -> 13 lessons on how to -> actually work with AI. -> It goes over things like -> effective prompting, critical -> thinking on the outputs, -> and it has a vocabulary -> sheet that you’ll want -> to read and save for later. -> And the third one is -> Intro to Claude Cowork -> Claude Cowork is where -> you can actually get stuff -> done with Claude. -> So, this course covers -> projects, skills, plug-ins, -> scheduling tasks, handling -> files, and then also how -> to pick the right model -> for the job, and then when -> you finish these courses, -> just go to your LinkedIn -> and go to your profile -> Click "Add section," and -> then go to "Licenses -> and certifications" and -> add all three of these. -> And then when you land -> the interview, you should -> talk about your AI fluency -> often as you possibly can -> I feel pretty confident that -> you’ll truly be able to -> differentiate yourself -> from other candidates -> if you do this

BeingInvested

13,130 görüntüleme • 3 ay önce

🚨🚨 Ed Elson says the government should subpoena OpenAI and Anthropic and demand the evidence behind their alarming claims: Elson: If Anthropic believes that they have created a technology that might destroy society, I'm waiting for Anthropic to come out there and say, here is the empirical evidence. As soon as an Anthropic employee comes out and says what he said, that is a corporate communication by the company. It was by an Anthropic employee. The entire organization is responsible for the words that were published and transmitted to literally millions of people all over the world. They are accountable for that. Now let's talk about the government side of this. The government is accountable to pursuing action against illegal activity. If someone is building a technology that is going to murder people or destroy society, I think we can all agree that that is a bad thing and it's illegal. You can't be doing that. That's now the government's responsibility to go out and pursue action against these companies. Now, what do you do if you want to pursue action? Well, it needs, again, we need to investigate their claims. The government should hand a subpoena to Anthropic and to OpenAI and say, “Hey, you've just said something very, very serious. We're going to take it seriously. We want you to show us the evidence and show us your claims.” And if it's true that what they've built is seriously bad, then we need to start working together on this. If it turns out that what they're saying is hyperbolic and it's not true at all and it's just some guy who went online and wanted to get clicks and drum up some hysteria and some drama, then we're talking about a different problem. Now we're talking about things like securities fraud. Now we're talking about things like public deception.

MeidasTouch

43,763 görüntüleme • 14 gün önce