Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

I've been trying Meta smart glasses' new multimodal AI - while it's pretty basic right now, it's still sick to see it combine what it sees from the camera with the language model to describe what it's seeing! Already solid for accessibility Full episode:

1,128,504 görüntüleme • 2 yıl önce •via X (Twitter)

10 Yorum

Ben Geskin profil fotoğrafı
Ben Geskin2 yıl önce

We are getting there 👀

Everett World profil fotoğrafı
Everett World2 yıl önce

Multi-modality is the next step. We're moving from LLMs to world models, that will be even more helpful for practical reasons.

Karthik Kannan profil fotoğrafı
Karthik Kannan2 yıl önce

Um, we’re already here

Average Engineer profil fotoğrafı
Average Engineer2 yıl önce

When he says that phrase and ask questions, Glass takes photo and takes question from users using OpenAI Whisper API, Upload to Gta-4 Vision API with your prompt Get back the results and make it speak again using API. Is there anything i am missing here?

David profil fotoğrafı
David2 yıl önce

This is the way Marques !

Rahul profil fotoğrafı
Rahul2 yıl önce

Lower hanging use cases I could use this for already: reading while walking.

Pawel profil fotoğrafı
Pawel2 yıl önce

It is going to be the feature of AR

Newtonian profil fotoğrafı
Newtonian2 yıl önce

Epic 👏🔥🔥

Joseph Bella profil fotoğrafı
Joseph Bella2 yıl önce

I feel like we are seeing history repeat itself when Apple released the Newton and Palm made the Pilot.

Ellie MacQueen profil fotoğrafı
Ellie MacQueen2 yıl önce

Here for the David stares

Benzer Videolar