Загрузка видео...

Не удалось загрузить видео

На главную

I've been trying Meta smart glasses' new multimodal AI - while it's pretty basic right now, it's still sick to see it combine what it sees from the camera with the language model to describe what it's seeing! Already solid for accessibility Full episode:

1,128,504 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 10

Фото профиля Ben Geskin
Ben Geskin2 лет назад

We are getting there 👀

Фото профиля Everett World
Everett World2 лет назад

Multi-modality is the next step. We're moving from LLMs to world models, that will be even more helpful for practical reasons.

Фото профиля Karthik Kannan
Karthik Kannan2 лет назад

Um, we’re already here

Фото профиля Average Engineer
Average Engineer2 лет назад

When he says that phrase and ask questions, Glass takes photo and takes question from users using OpenAI Whisper API, Upload to Gta-4 Vision API with your prompt Get back the results and make it speak again using API. Is there anything i am missing here?

Фото профиля David
David2 лет назад

This is the way Marques !

Фото профиля Rahul
Rahul2 лет назад

Lower hanging use cases I could use this for already: reading while walking.

Фото профиля Pawel
Pawel2 лет назад

It is going to be the feature of AR

Фото профиля Newtonian
Newtonian2 лет назад

Epic 👏🔥🔥

Фото профиля Joseph Bella
Joseph Bella2 лет назад

I feel like we are seeing history repeat itself when Apple released the Newton and Palm made the Pilot.

Фото профиля Ellie MacQueen
Ellie MacQueen2 лет назад

Here for the David stares

Похожие видео