Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

This week, grounding DINO 1.5 was released It is a new model that uses text prompts to detect objects from videos and images in real-time Examples & demo to try below:

56,027 Aufrufe • vor 2 Jahren •via X (Twitter)

10 Kommentare

Profilbild von Allen T.
Allen T.vor 2 Jahren

1) Video object detection

Profilbild von Allen T.
Allen T.vor 2 Jahren

2) Image object detection

Profilbild von Allen T.
Allen T.vor 2 Jahren

3) Paper: Website: Playground:

Profilbild von Allen T.
Allen T.vor 2 Jahren

4) Grounding DINO 1.5

Profilbild von LEYVERSE
LEYVERSEvor 2 Jahren

Someone, please combine this with Segment Anything and replace the boxes with masks of the objects. It would be incredible!

Profilbild von Allen T.
Allen T.vor 2 Jahren

I agree! That would be super helpful for editing

Profilbild von Dustin Hollywood
Dustin Hollywoodvor 2 Jahren

This would also rule out copyright violations because the law doesn’t distinguish between human and bot/code “seeing” and learning. What a great loop hole that was prolly not intended but awesome haha 😆

Profilbild von Allen T.
Allen T.vor 2 Jahren

I never even considered this angle, Dustin! 👀😯I wonder if that is going to be an argument that a lot of these companies that have been asking for permission to train on user phone and car data will use. It is allowing their AI to see the world without scraping

Profilbild von Happy
Happyvor 2 Jahren

wow

Profilbild von Allen T.
Allen T.vor 2 Jahren

Agree!

Ähnliche Videos