Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Last week we launched agentic video understanding in Gemini! It navigates timelines dynamically instead of ingesting every frame, cutting tokens by up to 88% and costs by up to 66%! it also supports YouTube links. you can test it in Google AI Studio (toggle it on in settings):

29,318 Aufrufe • vor 20 Tagen •via X (Twitter)

31 Kommentare

Profilbild von Patrick Loeber
Patrick Loebervor 20 Tagen

note that it is not enabled by default in the API. You can enable it by setting processing to "agentic". we also wrote a developer guide:

Profilbild von Shez Malik
Shez Malikvor 20 Tagen

@GoogleAIStudio gemini found the skip intro button

Profilbild von Erdem Demirci
Erdem Demircivor 20 Tagen

@GoogleAIStudio Works like magic! Tested it in AI Studio and it’s incredible. Will we be able to use this in AntiGravity as well?

Profilbild von Shubh
Shubhvor 20 Tagen

@GoogleAIStudio @grok does it still look at all the frames?

Profilbild von z
zvor 20 Tagen

@GoogleAIStudio Does Antigravity support video understanding (both video attachments and YouTube links)? Would be really cool

Profilbild von Stats Wire
Stats Wirevor 20 Tagen

@GoogleAIStudio I am testing using my youtube channel video link

Profilbild von a s
a svor 20 Tagen

@GoogleAIStudio I cant see where it is. Google always does a gazillion products no one knows off🤦‍♂️ no wonder most of them get shutdown

Profilbild von Tyson Hutchins
Tyson Hutchinsvor 20 Tagen

@GoogleAIStudio This available via the CLI as well?

Profilbild von Nikita
Nikitavor 20 Tagen

@GoogleAIStudio Add it in antigravity as well as aistudio pleeeeeeease

Profilbild von Alex Freitas
Alex Freitasvor 20 Tagen

@GoogleAIStudio That is a useful cost lever, but the benchmark should include retrieval misses and temporal jumps, not just token savings. If the model skips the decisive frame, an 88% reduction is a cheaper wrong answer.

Profilbild von Vladimir
Vladimirvor 20 Tagen

@GoogleAIStudio The agy harness and rules around using the subscription with other harness are doing a disservice to your model

Profilbild von White Us
White Usvor 20 Tagen

@GoogleAIStudio 88% fewer tokens but only 66% cheaper is the interesting part. the seeking itself costs something, and whatever tokens survive are the expensive ones

Profilbild von Armaan Singh
Armaan Singhvor 20 Tagen

The 66% cut is the headline. The open question is coverage. If the model skips a stretch, it can answer fluently without opening the decisive seconds. Fine for a creative prompt. Not fine for evidence or compliance. Wrote this up with the investing read (Alphabet vs Axon, Samsara, Motorola):

Profilbild von Flextor
Flextorvor 20 Tagen

@GoogleAIStudio Cutting 88% of tokens is the real flex. Frame by frame video analysis gets expensive fast.

Profilbild von ReidBKimball
ReidBKimballvor 20 Tagen

@GoogleAIStudio I love it! I'm building an app that creates linear issues from videos of usability tests for apps.

Profilbild von randomguy77
randomguy77vor 20 Tagen

@GoogleAIStudio add this to antigravity

Profilbild von Girish
Girishvor 20 Tagen

@GoogleAIStudio Thank you Gemini this is the one area of AI you are the undisputed king!

Profilbild von Ashkan
Ashkanvor 20 Tagen

@GoogleAIStudio I’m just waiting for Google to dominate AI in 2027, I have a feeling

Profilbild von blankbrain
blankbrainvor 20 Tagen

@GoogleAIStudio wow such a dumb approach

Profilbild von 野萌君Rumi
野萌君Rumivor 20 Tagen

@GoogleAIStudio Does the Google cloud agent platform support agentic video understanding?

Profilbild von Blue Akash
Blue Akashvor 20 Tagen

@GoogleAIStudio Can it also give the transceitp as it is?

Profilbild von S M
S Mvor 20 Tagen

@GoogleAIStudio Can you ask gemini for speech to text for video and then post it 😭

Profilbild von Sora
Soravor 20 Tagen

@GoogleAIStudio @patloeber @GoogleAIStudio Any issue with gemini-3.1-pro-preview? In AI Studio web (incl. Build) & Gemini API Tier 3, thinking and generation are faster, but quality is lower and words are fewer. Suspect output is routed to Lite. If it is a bug, please check and fix!

Profilbild von Biketommy
Biketommyvor 20 Tagen

@GoogleAIStudio 88% is significant success. And I think you can go beyond that.

Profilbild von Flextor
Flextorvor 20 Tagen

@GoogleAIStudio The token cut is the useful part. Skipping irrelevant frames makes long video analysis feel like retrieval instead of brute force.

Profilbild von Lite Casual
Lite Casualvor 20 Tagen

@GoogleAIStudio This is huge for robotics training

Profilbild von Sourav
Souravvor 20 Tagen

@GoogleAIStudio is this active in gemini/antigravity ?

Profilbild von Deon Holo
Deon Holovor 19 Tagen

@GoogleAIStudio Is the feature available in Antigravity IDE?

Profilbild von Gamma
Gammavor 20 Tagen

@GoogleAIStudio how about in Agy :D

Profilbild von Mauricio Solano 🇻🇦🇲🇽
Mauricio Solano 🇻🇦🇲🇽vor 20 Tagen

@GoogleAIStudio Is there a length cap? I have a 1 and 2 hour videos I wanna test

Profilbild von ali alshami
ali alshamivor 20 Tagen

@GoogleAIStudio Its way to expensive in api even in 5.3 flash lite Way too much output tokens

Ähnliche Videos