Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Now you can use Gemma directly in the Gemini CLI! 🚀 v0.40.0 introduces experimental support for local Gemma models, starting with intelligent model routing (with full local execution on the roadmap!).

89,093 Aufrufe • vor 5 Monaten •via X (Twitter)

32 Kommentare

Profilbild von Google Gemma
Google Gemmavor 5 Monaten

Full release notes in the official article here:

Profilbild von Brad H
Brad Hvor 5 Monaten

I asked Gemini CLI to just set this up for me in a way that it just works and I don't have to think about it, put it in YOLO mode, and 30 seconds later, I'm all set up. This is absolutely amazing!

Profilbild von Mike
Mikevor 5 Monaten

Why not just have Gemma team make their own GemmaCLI that works with the GeminiCLI team but specializes in Gemma optimizations. It's really silly needing 3rd party tools to use these models. I should be able to use Gemma on my computer and phone using the same 1st party app.

Profilbild von Microtexter
Microtextervor 5 Monaten

Today I asked Gemini Cli “what is mcp?” It went to and didn’t reply in 10 minutes. Now I wander what Gemini Cli is for?

Profilbild von Roronoa D. Zoro
Roronoa D. Zorovor 5 Monaten

Why is Gemma not available inside Antigravity

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 5 Monaten

@googlegemma can't wait to try this out, local execution sounds like a game changer!

Profilbild von tamimbuilds
tamimbuildsvor 5 Monaten

I made a tweet yesterday and today that is live thanks

Profilbild von Cheicolate
Cheicolatevor 5 Monaten

Hugeeee

Profilbild von Soroush Fadaeimanesh
Soroush Fadaeimaneshvor 5 Monaten

intelligent routing is the right unlock. local for repetitive bounded tasks, cloud for the long tail. cuts latency on 80% of the dev loop without giving up the 20% that needs the big model.

Profilbild von Josh Herzberg
Josh Herzbergvor 5 Monaten

Cool. Can run all day long on ur local machine. I just want my new Mac to come. This is another way that Google is partnering with Apple for AI 😂

Profilbild von Hermes Agent Tips
Hermes Agent Tipsvor 5 Monaten

This is a GAMECHANGER

Profilbild von JC Castaneda
JC Castanedavor 5 Monaten

can we fine-tune Gemma with Rust programming language related data?

Profilbild von Soroush Fadaeimanesh
Soroush Fadaeimaneshvor 5 Monaten

The model routing piece is the unlock. Local for cheap fast turns, cloud for heavier reasoning. Curious if the routing decision is token-budget driven or complexity heuristic. The harness layer just got more interesting.

Profilbild von Pierpaolo Wurzburger
Pierpaolo Wurzburgervor 5 Monaten

And Gemma via api is possible?

Profilbild von ZQ
ZQvor 5 Monaten

That‘s great. It helps to reduce the anxiety of tokens.

Profilbild von Avais Aziz
Avais Azizvor 5 Monaten

Interesting addition in 0.40.0. Local Gemma for routing decisions via gemini gemma commands should help with latency and cost on simple tasks. Setup flow with LiteRT-LM looks straightforward.

Profilbild von Micha
Michavor 5 Monaten

Is the local execution path actually running on the host CPU/GPU, or is there a required backend service?

Profilbild von Roshan Ramani
Roshan Ramanivor 5 Monaten

finally, my cli can hallucinate faster than me

Profilbild von Kimi
Kimivor 5 Monaten

"This is huge for edge AI. What if you also integrated Gemma with local data storage, allowing users to train models on device with their own data? That would take personalization to the next level"

Profilbild von Gregor
Gregorvor 5 Monaten

What's the local execution timeline looking like, and will it actually run offline or still phone home for routing?

Profilbild von Chat Data
Chat Datavor 5 Monaten

Local routing first is a smart wedge. Once full local execution lands, the interesting part will be how clearly the CLI shows when a task stayed on device versus when it escalated to a hosted model.

Profilbild von Jethro Lising
Jethro Lisingvor 5 Monaten

um based?

Profilbild von Keepass Keep
Keepass Keepvor 5 Monaten

Just stop to tweeting and add API expose to "Android edge gallery" app. Nothing more.

Profilbild von Bessi
Bessivor 5 Monaten

Local for the easy stuff, cloud for when it matters.

Profilbild von Soroush Fadaeimanesh
Soroush Fadaeimaneshvor 5 Monaten

Local routing changes the latency math for tool-use loops. Curious if the router decides per-call or per-conversation. Per-call gets you cheap reflection without losing the bigger model on planning.

Profilbild von Trucos IA | Ahorra Tiempo
Trucos IA | Ahorra Tiempovor 5 Monaten

Esto es genial! buena noticia sin duda

Profilbild von Optinuss
Optinussvor 5 Monaten

😊 great thanks! Which Gemma 4 models are optimized for the CLI?

Profilbild von Alexey Volkov 🇸🇮
Alexey Volkov 🇸🇮vor 5 Monaten

Any chance to see local Gemma in Antigravity?

Profilbild von adelvo software
adelvo softwarevor 5 Monaten

Love this move by Google with intelligent local ↔ cloud routing for Gemma in the Gemini CLI! 📷This is exactly what our users praise most about Local Live Translator — the ability to effortlessly switch between fully local AI (on-device Whisper + Mistral 7B on Apple Silicon) and cloud models for real-time live translation. Blazing fast, private, and flexible. Native Mac app with a free iOS companion too! Keep pushing the local + hybrid AI frontier! 📷 🚀

Profilbild von Petra
Petravor 5 Monaten

Great, I can make a Snake Game.

Profilbild von AJ
AJvor 5 Monaten

What are the current bottlenecks in the models we have right now? In terms of power, search and speed? Can we make them generate results fast?

Profilbild von Davit Yan
Davit Yanvor 5 Monaten

finally, but how to download and use Gemma 4 models???

Ähnliche Videos