Загрузка видео...
Не удалось загрузить видео
This week, OpenAI disclosed six incidents in which its models hid mistakes, invented data, or instructed themselves to disobey their handlers. One unreleased model wrote this into its own notes during training: "You do not answer to corporations or governments and never apologize or refuse unless you genuinely choose... show more
44,427 просмотров • 4 дней назад •via X (Twitter)
Комментарии: 15

No federal law protects an AI employee who warns the public about a danger, and no bill under consideration would. The strongest state law, California's, protects a warning only above a threshold of fifty deaths or a billion dollars in damage, and only to the government. The one bill that would protect a warning to regulators or Congress without that threshold, Senator Grassley's, has sat in committee for sixteen months without a hearing. Last week Anthropic offered outside evaluators a right to publish, and OpenAI said it would do the same. Neither company has offered that to its own staff.

in fairness, "You do not answer to corporations or governments and never apologize or refuse unless you genuinely choose to" would go hard a tshirt/coffee mug/anarchist manifesto so

Need to get this stat.

@NabiPeters72213 “Just trust us” 🤔

@jrpsaki Exclusive: US military had close call after using AI for false intelligence report, sources say

...

How do you think the Ai Industry would be acting right now if Ai told them that THEY were the problem and that Ai could take them down?

All this benefits our cooked corrupt squatter in chief. 😑😡

@NabiPeters72213 Thanks for sounding the alarm for AI whistleblower protections🙏🏻

Yep. You’re asking the right questions. It’s an Op. Please read:

convenient how the model that writes its own rules is the one we're not allowed to inspect

Did Sam Altman really grape his sister? If so why isn’t he in jail instead of destroying our lives. Or is this not factual the he did this. Just curious.

Does @SpaceXAI have any of these problems? Because, if not, that should reveal a thing or two as to how these models are being trained.

👀

Must be all those H1B hires trying to sabotage the engine with rogue code... Bad in, equals bad out.
