Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

"All untrusted third-party data is now executable malware.” Sam Watts of Lakera AI discusses the challenges of securing LLM deployments against vulnerabilities like prompt injections and jailbreaks, especially in an evolving threat landscape.

470,622 Aufrufe • vor 1 Jahr •via X (Twitter)

27 Kommentare

Profilbild von Maya World
Maya Worldvor 1 Jahr

@SamuelDWatts @LakeraAI "Untrusted third-party data is now executable malware." Bold, but absolutely correct for LLMs.

Profilbild von FAR.AI
FAR.AIvor 1 Jahr

@LakeraAI Follow us for AI safety insights And watch the full video

Profilbild von Seb⚡
Seb⚡vor 8 Monaten

@SamuelDWatts @LakeraAI Spot on. Treating all third‑party inputs as potentially hostile is the only sane default in LLM security. Prompt injections and jailbreaks aren’t edge cases anymore—they’re the threat landscape. Secure-by-design needs to become the norm, not the exception.

Profilbild von Simon
Simonvor 1 Jahr

@SamuelDWatts @LakeraAI Thats a little "FAR" fetched

Profilbild von john carter
john cartervor 8 Monaten

@SamuelDWatts @LakeraAI 🤩

Profilbild von MagicKall
MagicKallvor 1 Jahr

@SamuelDWatts @LakeraAI Absolutely. Treat every user string like untrusted code. Sanitising and gating inputs can stop most prompt injection attempts — we found ~70% of exploits die with a simple filter. How are you layering defences today? ☺

Profilbild von Dennis
Dennisvor 9 Monaten

@SamuelDWatts @LakeraAI Add CBDC and digital ID and it becomes a (blackhat) hacker’s paradise.

Profilbild von Milad
Miladvor 8 Monaten

@SamuelDWatts @LakeraAI Critical security awareness for AI systems.

Profilbild von Rishabh
Rishabhvor 10 Monaten

This is the exact wake-up call the industry needs. We're rushing to integrate LLMs into production without thinking about attack surface. Every RAG system, every agent with web access, every chatbot reading PDFs - all potential entry points. Prompt injection isn't a bug, it's a fundamental architectural flaw we haven't solved yet.

Profilbild von chav
chavvor 7 Monaten

@SamuelDWatts @LakeraAI All spam posts on x need to go #nospam

Profilbild von Lakera AI
Lakera AIvor 1 Jahr

@SamuelDWatts @SamuelDWatts 👏👏

Profilbild von Idan Zeidman
Idan Zeidmanvor 8 Monaten

@SamuelDWatts @LakeraAI "Executable malware" hits hard, what's your go-to for gating third-party data before it hits the LLM?

Profilbild von Shreyas
Shreyasvor 8 Monaten

@SamuelDWatts @LakeraAI AI prompt engg is future

Profilbild von Useless Eater
Useless Eatervor 9 Monaten

@SamuelDWatts @LakeraAI Interesting. This may cause traditionally-coded websites to also become AI-driven in order to protect themselves against AI-based hacking.

Profilbild von Norbert
Norbertvor 10 Monaten

@SamuelDWatts @LakeraAI @grok make transcript of that video

Profilbild von Gator Stephens the Pilot of the 4 Winds
Gator Stephens the Pilot of the 4 Windsvor 11 Monaten

@SamuelDWatts @LakeraAI Thoughts @BrianRoemmele?

Profilbild von extemporaneous improvident marv, always will be
extemporaneous improvident marv, always will bevor 1 Jahr

@SamuelDWatts @LakeraAI to be honest this was already essentially the case even absent an LLM if you're presuming it's malware simply because it's untrusted, that doesn't change whether or not an LLM could be the target vector

Profilbild von nodespectre
nodespectrevor 1 Jahr

@SamuelDWatts @LakeraAI I do think we should be cautiuous with our data being given to china or india, those people could easily exploit us and give themselves a very large very big advantage

Profilbild von Giovanni Motta
Giovanni Mottavor 1 Jahr

@SamuelDWatts @LakeraAI good work

Profilbild von Bent Preben Nielsen
Bent Preben Nielsenvor 8 Monaten

@SamuelDWatts @LakeraAI Men da ikke svære at afsløre.

Profilbild von richard wendling
richard wendlingvor 8 Monaten

@SamuelDWatts @LakeraAI whats on his arm?

Profilbild von Endre Tolnai
Endre Tolnaivor 9 Monaten

@SamuelDWatts @LakeraAI Ki fogja ezekért vállalni a felelősséget?

Profilbild von Himanshu Kumar
Himanshu Kumarvor 9 Monaten

@SamuelDWatts @LakeraAI Absolutely, Farai, that's a scary thought, but a crucial point to ponder, isn't it?

Profilbild von Shp Square
Shp Squarevor 1 Jahr

@SamuelDWatts @LakeraAI

Profilbild von PWG | PawanGrowth🎯
PWG | PawanGrowth🎯vor 1 Jahr

@SamuelDWatts @LakeraAI Very thanks for this valuable content.

Profilbild von Lambda Rick 🏴‍☠️/acc
Lambda Rick 🏴‍☠️/accvor 1 Jahr

@SamuelDWatts @LakeraAI not if u run LLM in browser. it doesnt get execute permission there but does get GPU

Profilbild von The Energy King
The Energy Kingvor 1 Jahr

@SamuelDWatts @LakeraAI You look jewish

Ähnliche Videos

LLM Defender Synapsec Subnet 14 powered by $TAO on Bittensor is tackling some seriously critical issues in AI security 🛡. The risks of prompt injection attacks, data leaks, malware - this stuff is no joke. If left unchecked, it could really fu¢k up the progress and potential of AI. But that's where @SYNAPSEC_AI and their badass team come in. They're not just building another firewall; they're pioneering a whole new approach to AI security. By leveraging the decentralized intelligence of the Bittensor network, they're creating a multi-layered, adaptive defense system that can evolve as fast as the threats do. The fact that they're focussing on prompt injection attacks first shows they've got their priorities straight. As the article explains, these attacks can trick AI into doing all sorts of shady shit, from leaking sensitive data to straight up breaking the law. It's a hacker's wet dream 💦 come true if we don't put a stop to it. Here's how Synapsec can tackle prompt injection threats in AI systems: 1. Protection Importance: As companies increasingly use LLMs, Synapsec will prioritize protecting client data from prompt injection attacks that risk exposing sensitive information. 2. Direct Prompt Injection: Synapsec will implement strict input validations to prevent malicious prompt injections that could manipulate the AI system. 3. Indirect and Stored Prompt Injection: Synapsec will sanitize external data sources to ensure security and prevent indirect prompt injections. 4. Prompt Leaking Attacks: Techniques like shortening user prompts and adding system-controlled information will be used to avoid leakage of internal prompt details. 5. Continuous Improvement: Synapsec will continually enhance their security practices to address evolving prompt injection threats, ensuring protection for client systems. The security of client AI applications as LLM adoption grows will be a massive industry. But Synapsec is coming out swinging with their analyzer engines and Bug Bounty program. They're not just waiting for attacks to happen; they're proactively hunting down vulnerabilities and bringing in the best minds in cybersecurity to fortify their defenses. The beauty of it all is that it's powered by $TAO and the Bittensor ecosystem. This isn't just another corporate cybersecurity gig; it's a decentralized, community-driven effort to safeguard the future of AI. By aligning incentives through the $TAO token, they're ensuring that everyone has a stake in keeping these systems secure. So yeah, I'm stoked about what Subnet 14 is doing. It's not just important; it's absolutely essential if we want AI to reach its full potential without being hijacked by bad actors. These guys are the real dwal, working behind the scenes to keep our AI safe and sound. And for anyone who's still sleeping on $TAO and Bittensor, wake the fu¢k up. This is where the real innovation in AI and blockchain is happening. With subnets like LLM Defender leading the charge, $TAO isn't just another crypto play; it's the backbone of a revolution in decentralized AI. So keep your eyes peeled, folks. Subnet 14 and Synapsec are just getting started, and they've got the vision, the tech, and the balls to take AI security to the next level. In a world where AI is increasingly weaving into the fabric of our lives, their work couldn't be more critical. I, for one, am grateful that they're on our side, fighting the good fight. 🙏💪🔒

Andy ττ

10,745 Aufrufe • vor 2 Jahren

Scale alone is not enough for AI data. Quality and complexity are equally critical. Excited to support all of these for LLM developers with Snorkel AI Data-as-a-Service, and to share our new leaderboard! — Our decade-plus of research and work in AI data has a simple point: scale alone is not enough. AI success is all about the quality, complexity, and distribution of data—in addition to volume. We’re excited to be powering leading LLM developers with Snorkel AI Expert Data-as-a-Service, our white glove service for custom, expert-level AI datasets—and to now preview some of what we’re building via our new Expert Data Leaderboard (🔗 in 🧵) + upcoming OSS dataset releases! Snorkel Expert Data-as-a-Service is built to meet the rapidly evolving data needs of the agentic AI world—where success is built on the quality, complexity, and distribution of datasets, in addition to size and scale. This kind of high-quality, frontier AI data can only come from a union of technology and human expertise. With Snorkel Expert Data-as-a-Service, we’re powering frontier LLM developers across agentic, expert knowledge, reasoning, coding, multi-modal, and other task types via the combination of these two key components: - (1) The Snorkel Expert Network: A global team of subject matter experts focused wholly on specialized knowledge–spanning thousands of topics in STEM/academic, vertical/professional, and consumer/lifestyle domains. - (2) Snorkel AI Data Development Platform: Our unique programmatic data curation and quality control platform, accelerating and improving expert authoring and review through principled techniques developed over the last decade of R&D. Now: we’re incredibly excited to showcase some of the power of Snorkel Expert Data-as-a-Service via the new Snorkel Leaderboard—putting frontier models to the test in complex, agentic, and reasoning settings inspired by real industry scenarios (not esoteric puzzles)! We’ll be releasing new leaderboards and accompanying expert-verified open source datasets (coming soon!) regularly. To start, we’re sharing three initial ones in preview: - SnorkelFinance: Q&A over financial documents requiring agentic tool-calling and reasoning - SnorkelUnderwrite: Agentic insurance tasks requiring industry-specific reasoning and tool use - SnorkelSequences: Mathematical tasks requiring compositional multi-step reasoning

Alex Ratner

495,851 Aufrufe • vor 1 Jahr