Загрузка видео...

Не удалось загрузить видео

На главную

Introducing GPT-4o, our new model which can reason across text, audio, and video in real time. It's extremely versatile, fun to play with, and is a step towards a much more natural form of human-computer interaction (and even human-computer-computer interaction):

4,362,664 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 9

Фото профиля Greg Brockman
Greg Brockman2 лет назад

The new Voice Mode will be coming to ChatGPT Plus in upcoming weeks.

Фото профиля Greg Brockman
Greg Brockman2 лет назад

GPT-4o can also generate any combination of audio, text, and image outputs, which leads to interesting new capabilities we are still exploring. See e.g. the "Explorations of capabilities" section in our launch blog post ( or these generated images:

Фото профиля Greg Brockman
Greg Brockman2 лет назад

We also have significantly improved non-English language performance quite a lot, including improving the tokenizer to better compress many of them:

Фото профиля Dennis
Dennis2 лет назад

Dozens of startups obliterated

Фото профиля Benjamin BLM
Benjamin BLM2 лет назад

Audio, text and image? What's in the training data? You need literal billions of images and text tokens. Where did you get them, from the internet?

Фото профиля AshutoshShrivastava
AshutoshShrivastava2 лет назад

Desktop app with Vison is the best feature which was launched definitely game changer.

Фото профиля Charlene Wang
Charlene Wang2 лет назад

gpt4-o’s real time translation is gonna go viral & support the best consumer AI hardware. Saw it translate group conversations and respond in different language. It’s super fast and I can’t wait to use it in Japan!

Фото профиля illusion diffusion
illusion diffusion2 лет назад

where?

Фото профиля Alex Sharp
Alex Sharp2 лет назад

now that ai's can talk to each other i can finally delete all in from my podcast player

Похожие видео