Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

"Please learn from our mistakes. Don't do exactly the same things that we did, or you'll end up in ten years with having nothing to show for it." — Nicholas Carlini urging AI researchers to avoid the pitfalls of past adversarial ML research at the Vienna Alignment Workshop 2024.

5,370,650 görüntüleme • 1 yıl önce •via X (Twitter)

10 Yorum

FAR.AI profil fotoğrafı
FAR.AI1 yıl önce

Follow us for updates about upcoming content and workshops: and watch the full video at

Nilesh Trivedi profil fotoğrafı
Nilesh Trivedi1 yıl önce

Sounds good if the problem you're trying to solve is "how to publish 9000 papers in 10 years". 😁

floating point profil fotoğrafı
floating point1 yıl önce

I actually listened to the talk and I am not sure what’s the advice? It appears the advice is to study why adversarial papers are still so easy to break and maybe try not to put such a difficult problem as a target? Eg limit jailbreak to 100 queries via API (no weights access).

Janosch profil fotoğrafı
Janosch1 yıl önce

One of the first things I used Chat-GPT for was having it write for me a Python keylogger. I just had to avoid that word and prompt around it. The second thing I did was having it write porn stories. Worked incredibly well, with some understanding of context windows.

Noir Tintin - profil fotoğrafı
Noir Tintin -1 yıl önce

He looks like Andrew tate if he went to school

Kass Popper profil fotoğrafı
Kass Popper1 yıl önce

Alignment evals for LLMs based purely on input-output pattern recognition are bound to fail.

Gianluca profil fotoğrafı
Gianluca1 yıl önce

Making mistake it's a fundamental part of human nature, if you don't make mistakes you don't progress,it's necessary and it's thanks to the mistakes of the past that progress was made possible. If you eliminate mistake, you cut a central part of human nature.

Tensor Templar profil fotoğrafı
Tensor Templar1 yıl önce

Oh someone elaborating on negative results for once, cool.

Burek profil fotoğrafı
Burek1 yıl önce

Top G looks different

Revised Basics profil fotoğrafı
Revised Basics1 yıl önce

These AI researchers are just working together to create more recessions and crisis for middle and lower class.

Benzer Videolar