Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🕳️🐇Into the Rabbit Hull – Part I (Part II tomorrow) An interpretability deep dive into DINOv2, one of vision’s most important foundation models. And today is Part I, buckle up, we're exploring some of its most charming features.

64,170 görüntüleme • 11 ay önce •via X (Twitter)

19 Yorum

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Assuming the Linear Rep. Hypothesis, SAEs arise naturally as instruments for concept extraction, they will be our companions in this descent. Archetypal SAE uncovered 32k concepts. Our first observation: different tasks recruit distinct regions of this conceptual space.

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Let's zoom in on classification. For every class, we find two concepts: one fires on the object (e.g., "rabbit"), and another fires everywhere *except* the object -- but only when it's present! We call them Elsewhere Concepts (credit: @davidbau).

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

This kind of concept breaks a key assumption in interpretability: that a concept is about the tokens where it fires. Here it is the opposite—the concept is defined by where it does not fire. An open question is how models form such concepts.

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Another surprise here: the most important concepts are not object-centric at all, but boundary detectors. Remarkably, these concepts coalesce into a low-dimensional subspace within (see paper).

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Now for depth estimation. How does DINO know depth? It turns out it has discovered several human-like monocular depth cues: texture gradients resembling blurring or bokeh, shadow detectors, and projective cues. Most units mix cues, but a few remain remarkably pure.

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Curious tokens, the registers. DINO seems to use them to encode global invariants: we find concepts (directions) that fire exclusively (!) on registers. Example of such concepts include motion blur detector and style (game screenshots, drawings, paintings, warped images...)

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Huge thanks to all collaborators who made this work possible, and especially to @WangBinxu. This work grew from a year of collaboration! Tomorrow, Part II: geometry of concepts and Minkowski Representation Hypothesis. 🕹️ 📄

Sebastian A. profil fotoğrafı
Sebastian A.11 ay önce

@Napoolar Wooow amazing work! is it possible to apply it with a fine tuned model?

Nikhil Parthasarathy profil fotoğrafı
Nikhil Parthasarathy11 ay önce

@Napoolar @akjags Awesome stuff!

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

@akjags Thx Nikhil ! Same goes for you 😉

Kosta Derpanis at #ECCV2026 🇸🇪 profil fotoğrafı
Kosta Derpanis at #ECCV2026 🇸🇪11 ay önce

@Napoolar 💪

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Thx Kosta 😉 !!

Akshay Jagadeesh profil fotoğrafı
Akshay Jagadeesh11 ay önce

@Napoolar Congrats Thomas! Incredible work!!!

Thomas Fel profil fotoğrafı
Thomas Fel11 ay önce

Thx Akshay 🤠

Ali Shehral profil fotoğrafı
Ali Shehral11 ay önce

@Napoolar Amazing work, @Napoolar! Looking forward to diving into the paper this weekend!

Dustin profil fotoğrafı
Dustin11 ay önce

@Napoolar Looking forward to Part II! It's exciting to see such an in-depth exploration of DINOv2’s capabilities and interpretability—can’t wait to discover what insights you’ll share next.

Agam Chaudhary profil fotoğrafı
Agam Chaudhary11 ay önce

@Napoolar Looking forward to this series. DINOv2 remains such a fascinating model for how it captures visual structure without supervision. What part of its interpretability are you most curious to unpack first—feature emergence or layer level abstraction?

Out of service profil fotoğrafı
Out of service11 ay önce

@Napoolar Fascinating work.

StupidFood profil fotoğrafı
StupidFood11 ay önce

@Napoolar "Into the Rabbit Hull"? My therapist is going to have *so* many questions. Buckled up and praying for a smooth landing, not another existential crisis. 🐰✨

Benzer Videolar