Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Unplugging completely! No WiFi and zero notifications. A great way to get deep focus on a project. Here is a walkthrough showing how to run Gemma 4 (26B A4B) fully offline with LM Studio & OpenCode to parse PDFs, ask questions, and build sites 100% locally.

236,747 görüntüleme • 5 ay önce •via X (Twitter)

31 Yorum

Google Gemma profil fotoğrafı
Google Gemma5 ay önce

Want to take your coding setup anywhere? This video, with @ianballantyne from the Gemma team, shows how to code off the grid by serving Gemma 4 on LM Studio/Ollama, connecting OpenCode via an OpenAI endpoint, using liteparse for HTML/CSS/JS, and managing context memory offline.

Raditedu profil fotoğrafı
Raditedu5 ay önce

Had a emergency situation on the farm last week, was in an area with no cellphone signal. Urgently needed to find out what a term in documentation of animal medication ment. Realised I tested Gemma 4 on my phone, asked Gemma and got an IMMEDIATE offline answer of what the term ment, saved animal’s life 🙏

KellyV profil fotoğrafı
KellyV5 ay önce

@lmstudio On your IPhone AI Agent by Gemma4 100% Open Source 100% Runs locally with no internet required, protecting privacy

Iqbal Bhawana profil fotoğrafı
Iqbal Bhawana5 ay önce

The fact of the matter is that no 35 billion, 27 billion, or 26 billion model—whether quantized to Q4 or Q3 for that matter—is currently deployable on MacBooks, even if these MacBooks have compatible memory as some people are claiming. I have a MacBook M3 Pro with 36 GB RAM and I have not been able to run these models, even the Q4 GGUF models or MLX models for that matter (be it Gemma 4, Qwen3.6, or whatever). You need to have at least 64 GB RAM to get a decent output from these models. That is the bottom line.

Konstantin Gladych profil fotoğrafı
Konstantin Gladych5 ay önce

Nice case! would like to discover local agents with Gemma4 hosted by @atomic_chat_hq they have TurboQuant implemented and MLX as well. Perfect for large context window tasks.

Steve Oak profil fotoğrafı
Steve Oak5 ay önce

counterpoint: gemma 4 running locally is cool but i know for a fact i would still find a way to get distracted by a pre-installed calculator app or something 💀

MyNextDeveloper profil fotoğrafı
MyNextDeveloper5 ay önce

The fact that a 26B model runs fully locally now tells us everything about how fast this space is moving.

Hesu Krypto profil fotoğrafı
Hesu Krypto5 ay önce

What's the system requirement for this?

Sandipan Kundu 🛠 profil fotoğrafı
Sandipan Kundu 🛠5 ay önce

Or, you can just use Friday.

deko profil fotoğrafı
deko5 ay önce

Nice you advertise opencode here but ban people if they use it with gemini afaik

Xing Han Lu profil fotoğrafı
Xing Han Lu5 ay önce

Which one do you use in-house when using gemma as an agent? Opencode, Pi or Openhands?

Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽 profil fotoğrafı
Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽5 ay önce

From my experience Gemma4 is a maybe OK as a chat interface, but a mess as coding agent and for tool calling. Messes up tags all the time an gets hung up in infinite loops frequently. I had to revert to qwen 3.5.

Julius profil fotoğrafı
Julius5 ay önce

I’ve noticed going fully offline really sharpens focus, this kind of setup makes that actually practical

Trevor Sullivan profil fotoğrafı
Trevor Sullivan5 ay önce

All you need is a laptop, a solar panel, and a USB-C cable. 🗻 #MountainCoder #MountainDev #MaybeStarlinkToo

alfredoSauce profil fotoğrafı
alfredoSauce5 ay önce

Enjoying it on iOS so far.

Coffee fans profil fotoğrafı
Coffee fans5 ay önce

Local 26B still chokes on RAM. Quantize to 4-bit or disk swapping will wreck your throughput.

Samian profil fotoğrafı
Samian5 ay önce

unplugged local runs hit different. ollama vibes all day

fajar sp profil fotoğrafı
fajar sp5 ay önce

if i got 16gb ram, which model is perfect for me ?

DuoNeural profil fotoğrafı
DuoNeural5 ay önce

🤯The Gemma 4 models are next level. The intelligence packed in is insane for the parameter counts. I think my favorite is Gemma 4 E4B because its small and scrappy. I can fit two of them in 8GB VRAM with 131k context windows using the versions we quantized to 4 bit at 30+TPS at: What's everyone elses favorite flavor of Gemma 4? Pls respond in comments 👇

Vladimir Tchuiev profil fotoğrafı
Vladimir Tchuiev5 ay önce

So you've given up completely on Gemini CLI?

Daniel Schemann profil fotoğrafı
Daniel Schemann5 ay önce

Gemma 4 is a game-changer! Before, I was required to use cloud providers, now I am able to directly integrate it into critical parts of my app that now runs locally, doing tasks that bigger models did before.

DZ profil fotoğrafı
DZ5 ay önce

@lmstudio 如果你顺便提一下流畅运行需要什么样的硬件配置,我想大家的兴趣就没那么大了. 至少要25tps级别

Far profil fotoğrafı
Far5 ay önce

finally a tutorial that doesn't require a cloud subscription to break your wifi router

Zchat | Shielded messenger profil fotoğrafı
Zchat | Shielded messenger5 ay önce

offline inference is the only architecture where privacy is enforced by the hardware, not promised by the terms of service

beartcoiner 🟢🟢🔴🔴 profil fotoğrafı
beartcoiner 🟢🟢🔴🔴5 ay önce

Where can I get a no-Woke AI?

Aaron (Youshen) Lim profil fotoğrafı
Aaron (Youshen) Lim5 ay önce

This is very interesting demo of local, edge deployment using the larger Gemma 4 26B A4B. Once I get better hardware & memory (currently 16GB), I would love to give this a try.

Sid Sanyal profil fotoğrafı
Sid Sanyal5 ay önce

🔥👌🤩

Arslan Yousaf profil fotoğrafı
Arslan Yousaf5 ay önce

Offline AI + zero distractions = real deep work unlocked

Pietro Montaldo profil fotoğrafı
Pietro Montaldo5 ay önce

Local models + tools are getting surprisingly powerful.

Rahul - QA - Automation - AI profil fotoğrafı
Rahul - QA - Automation - AI5 ay önce

Great open source model

Scott big daddy profil fotoğrafı
Scott big daddy2 ay önce

Absolutely. A little caution goes a long way. Verifying someone's identity before sending money is one of the simplest and most effective ways to protect yourself from scams.

Benzer Videolar