Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

DeepSeek-R1 crafted a jailbreak for itself that also worked for other AI models. Siva Reddy: R1 "complies a lot" with dangerous requests directly. When creating jailbreaks: long prompts, high success rate, "chemistry educator" = universal trigger. 👇

1,293,254 Aufrufe • vor 1 Jahr •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

U.S. Navy Bans DeepSeek Over 'Security Concerns' As 'Substantial' Evidence Emerges Chinese AI Ripped Off ChatGPT | ZeroHedge The U.S. Navy has instructed service members to avoid using the Chinese AI platform DeepSeek, citing "potential security and ethical concerns," according to CNBC. An email sent to "shipmates" in recent days, confirmed by CNBC on Tuesday, referenced the Navy's AI policy and emphasized the importance of refraining from using DeepSeek. The memo warned service members against using the platform "for any work-related tasks or personal use" and instructed them to "avoid downloading, installing, or using the DeepSeek model in any capacity." The warning follows the recent rise of DeepSeek’s R1 model, which has garnered significant attention worldwide, particularly within the U.S. business and technology sectors. The R1 model has demonstrated capabilities comparable to OpenAI’s models. In December, DeepSeek claimed it had successfully trained a large language model in just two months at a cost of $6 million—a figure disputed by technologists—despite U.S. restrictions on semiconductor chip exports to China. The R1, an open-source model, surged to the top of Apple’s app store rankings this week, triggering a market sell-off. Shares of AI chipmakers Nvidia and Broadcom plummeted by 17% on Monday, wiping out a combined $800 billion in market value. Nvidia has since recovered some of its losses. On Monday, DeepSeek announced a temporary restriction on user registrations, citing "large-scale malicious attacks" on its services, before later restoring normal operations. DeepSeek’s advancements have challenged the long-held belief that the U.S. was significantly ahead of China in AI development. Asked how R1 caught up to ChatGPT, AI and Crypto Czar David Sacks suggested that DeepSeek may have leveraged a technique known as "distillation" to train its model using OpenAI’s technology. “There’s a technique in AI called distillation, which you’re going to hear a lot about. It’s when one model learns from another model,” Sacks explained to Fox News. “Effectively, the student model asks the parent model millions of questions, mimicking the reasoning process and absorbing knowledge.” “They can essentially extract the knowledge out of the model,” he continued. “There’s substantial evidence that what DeepSeek did here was distill knowledge from OpenAI’s models.” “I don’t think OpenAI is too happy about this,” Sacks added. President Donald Trump has said that DeepSeek “should be a wake-up call” for U.S. tech companies. “The release of DeepSeek AI from a Chinese company should be a wake-up call for our industries that we need to be laser focused on competing,” the president told reporters ahead of a planned speech before Republican lawmakers in Florida. Read more:

Owen Gregorian

75,351 Aufrufe • vor 1 Jahr

Because you guys loved the 20 minutes of me asking the Humane Ai Pin voice questions so much, here's 19 minutes (almost 20!), no cuts, of me asking the rabbit inc. R1 AI questions and using its computer vision to "look" at stuff Some quick thoughts on the R1: • The AI/LLM is not perfect, but it gets way more correct than it does incorrect • The R1 is FAST to respond with answers. The Ai Pin looks embarrassingly slow in comparison • Vision is very impressive. Sometimes it IDs objects incorrectly (like in my other video, but since hard resetting, it seems to get more things right). It's also fast like the audio responses • Summaries are on point. I pointed the R1 at various Inverse articles that I either wrote or edited and it did a great job giving me the main points, even when the text was friggin' tiny on my iMac • The LLM is far more intelligent than on Ai Pin. It's better at understanding follow-ups with natural language. The Ai Pin is supposed to be contextual, but it often doesn't seem to remember what I said right before • It's late (3:30 am right now) so I have not connected my R1 to services like Spotify, Uber, or Midjourney. Will do that in the morning after I get some sleep. I'm very excited to see how Large Action Model (LAM) works and to teach the R1 to do stuff for me • There are some bugs that and Peiyu Liao tell me they're working on. For example, fixing the time (very important) and adding the % symbol back (also very important if your Wi-Fi password uses it!). Somehow, they seem to be working faster to fix bugs and issues than Humane • Jesse also tells me they're paying close attention to feedback on the sensitivity of the analog scroll wheel. It doesn't feel responsive enough, sometimes lagging a half second behind your actual scroll. He says they tuned it to be less sensitive to prevent it from activating on surfaces like a table. I think it could be a smidge more responsive. At least, that can be adjusted in a future software update This is not a review, only first impressions. I need to actually spend real time using and, most importantly, living with the R1. That being said, my initial setup bugginess/issues aside, the R1 is (as you can see in this long video) working quite well. Again, not perfectly every time, but far better than the Ai Pin. I am impressed. Really, really impressed Drop your questions and I'll answer them in the morning. What an exciting moment in tech. I live for this kinda stuff!

Ray Wong

721,838 Aufrufe • vor 2 Jahren