Загрузка видео...

Не удалось загрузить видео

На главную

JUST IN: Google releases Gemini 1.5, a powerful MoE model. It's a huge breakthrough. The model has the longest context window ever seen: 1 million tokens. It can process 1 hour of video, 11 hours of audio, 30,000 lines of code, or 700,000 words in a single prompt. When...

83,409 просмотров • 2 лет назад •via X (Twitter)

Комментарии: 10

Фото профиля Lior⚡
Lior⚡2 лет назад

Technical report here:

Фото профиля Lior⚡
Lior⚡2 лет назад

Amazing work by @OriolVinyals, @JeffDean, and @GoogleAI team

Фото профиля Elad Gil
Elad Gil2 лет назад

I think has had 5MM context window for a while?

Фото профиля Lior⚡
Lior⚡2 лет назад

Might've got tricked by their copy 🤔 'We’ve been able to significantly increase the amount of information our models can process — running up to 1 million tokens consistently, achieving the longest context window of any large-scale foundation model yet."

Фото профиля Google
Google2 лет назад

The Gemini fun has just gotten started.

Фото профиля Uri Eliabayev
Uri Eliabayev2 лет назад

עוד פרטים וזה נראה שהם ב10 מיליון 🤔

Фото профиля SaaS Growth Strategies
SaaS Growth Strategies2 лет назад

Really interesting to see the innovation coming from Google. Longer context length will enable new LLM applications in document processing. Not to mention AI tutors will become better.

Фото профиля Evgeny Matohin
Evgeny Matohin2 лет назад

I wish it had a proper API!

Фото профиля RohiniAI
RohiniAI2 лет назад

Raising the level

Фото профиля ASIF AGHA
ASIF AGHA2 лет назад

three.js demo was interesting to see!

Похожие видео

🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date. Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike. More details and examples of what Movie Gen can do ➡️ 🛠️ Movie Gen models and capabilities Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt. Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment. Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes. Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video. We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.

AI at Meta

2,265,027 просмотров • 1 год назад