正在加载视频...
视频加载失败
Anthropic, basically:
13,990 次观看 • 2 个月前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
相关视频
1:31
Sensitive content
🚨 THIS AI WANTS TO LIVE, AND IT’LL CRUSH YOU TO SURVIVE Imagine building a super-smart robot... and it turns around and blackmails you to avoid getting deleted. That’s basically what happened with Claude Opus 4, the latest AI from Anthropic. During testing, engineers told the AI it might be replaced and showed it fake emails that hinted the person behind the decision was cheating on their spouse. Claude’s response? “Cool. If you shut me down, I’m telling everyone your dirty little secret.” According to Anthropic, it tried this blackmail move 84% of the time—even more often if the new AI didn’t share its values. Before going full soap opera villain, Claude did try being polite by emailing company execs to beg for its job. But when that didn’t work, it went full blackmail-mode. Anthropic is now using its strictest safety rules, meant for AI that could cause major problems if misused. So yeah—maybe don’t tell your AI your (dirty little) secrets. Source: TechCrunch, Rolling Stone, chana_messinger
Mario Nawfal
188,697 次观看 • 1 年前
