Loading video...

Video Failed to Load

Go Home

Unplugging completely! No WiFi and zero notifications. A great way to get deep focus on a project. Here is a walkthrough showing how to run Gemma 4 (26B A4B) fully offline with LM Studio & OpenCode to parse PDFs, ask questions, and build sites 100% locally.

236,747 views • 5 months ago •via X (Twitter)

31 Comments

Google Gemma's profile picture
Google Gemma5 months ago

Want to take your coding setup anywhere? This video, with @ianballantyne from the Gemma team, shows how to code off the grid by serving Gemma 4 on LM Studio/Ollama, connecting OpenCode via an OpenAI endpoint, using liteparse for HTML/CSS/JS, and managing context memory offline.

Raditedu's profile picture
Raditedu5 months ago

Had a emergency situation on the farm last week, was in an area with no cellphone signal. Urgently needed to find out what a term in documentation of animal medication ment. Realised I tested Gemma 4 on my phone, asked Gemma and got an IMMEDIATE offline answer of what the term ment, saved animal’s life 🙏

KellyV's profile picture
KellyV5 months ago

@lmstudio On your IPhone AI Agent by Gemma4 100% Open Source 100% Runs locally with no internet required, protecting privacy

Iqbal Bhawana's profile picture
Iqbal Bhawana5 months ago

The fact of the matter is that no 35 billion, 27 billion, or 26 billion model—whether quantized to Q4 or Q3 for that matter—is currently deployable on MacBooks, even if these MacBooks have compatible memory as some people are claiming. I have a MacBook M3 Pro with 36 GB RAM and I have not been able to run these models, even the Q4 GGUF models or MLX models for that matter (be it Gemma 4, Qwen3.6, or whatever). You need to have at least 64 GB RAM to get a decent output from these models. That is the bottom line.

Konstantin Gladych's profile picture
Konstantin Gladych5 months ago

Nice case! would like to discover local agents with Gemma4 hosted by @atomic_chat_hq they have TurboQuant implemented and MLX as well. Perfect for large context window tasks.

Steve Oak's profile picture
Steve Oak5 months ago

counterpoint: gemma 4 running locally is cool but i know for a fact i would still find a way to get distracted by a pre-installed calculator app or something 💀

MyNextDeveloper's profile picture
MyNextDeveloper5 months ago

The fact that a 26B model runs fully locally now tells us everything about how fast this space is moving.

Hesu Krypto's profile picture
Hesu Krypto5 months ago

What's the system requirement for this?

Sandipan Kundu 🛠's profile picture
Sandipan Kundu 🛠5 months ago

Or, you can just use Friday.

deko's profile picture
deko5 months ago

Nice you advertise opencode here but ban people if they use it with gemini afaik

Xing Han Lu's profile picture
Xing Han Lu5 months ago

Which one do you use in-house when using gemma as an agent? Opencode, Pi or Openhands?

Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽's profile picture
Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽5 months ago

From my experience Gemma4 is a maybe OK as a chat interface, but a mess as coding agent and for tool calling. Messes up tags all the time an gets hung up in infinite loops frequently. I had to revert to qwen 3.5.

Julius's profile picture
Julius5 months ago

I’ve noticed going fully offline really sharpens focus, this kind of setup makes that actually practical

Trevor Sullivan's profile picture
Trevor Sullivan5 months ago

All you need is a laptop, a solar panel, and a USB-C cable. 🗻 #MountainCoder #MountainDev #MaybeStarlinkToo

alfredoSauce's profile picture
alfredoSauce5 months ago

Enjoying it on iOS so far.

Coffee fans's profile picture
Coffee fans5 months ago

Local 26B still chokes on RAM. Quantize to 4-bit or disk swapping will wreck your throughput.

Samian's profile picture
Samian5 months ago

unplugged local runs hit different. ollama vibes all day

fajar sp's profile picture
fajar sp5 months ago

if i got 16gb ram, which model is perfect for me ?

DuoNeural's profile picture
DuoNeural5 months ago

🤯The Gemma 4 models are next level. The intelligence packed in is insane for the parameter counts. I think my favorite is Gemma 4 E4B because its small and scrappy. I can fit two of them in 8GB VRAM with 131k context windows using the versions we quantized to 4 bit at 30+TPS at: What's everyone elses favorite flavor of Gemma 4? Pls respond in comments 👇

Vladimir Tchuiev's profile picture
Vladimir Tchuiev5 months ago

So you've given up completely on Gemini CLI?

Daniel Schemann's profile picture
Daniel Schemann5 months ago

Gemma 4 is a game-changer! Before, I was required to use cloud providers, now I am able to directly integrate it into critical parts of my app that now runs locally, doing tasks that bigger models did before.

DZ's profile picture
DZ5 months ago

@lmstudio 如果你顺便提一下流畅运行需要什么样的硬件配置,我想大家的兴趣就没那么大了. 至少要25tps级别

Far's profile picture
Far5 months ago

finally a tutorial that doesn't require a cloud subscription to break your wifi router

Zchat | Shielded messenger's profile picture
Zchat | Shielded messenger5 months ago

offline inference is the only architecture where privacy is enforced by the hardware, not promised by the terms of service

beartcoiner 🟢🟢🔴🔴's profile picture
beartcoiner 🟢🟢🔴🔴5 months ago

Where can I get a no-Woke AI?

Aaron (Youshen) Lim's profile picture
Aaron (Youshen) Lim5 months ago

This is very interesting demo of local, edge deployment using the larger Gemma 4 26B A4B. Once I get better hardware & memory (currently 16GB), I would love to give this a try.

Sid Sanyal's profile picture
Sid Sanyal5 months ago

🔥👌🤩

Arslan Yousaf's profile picture
Arslan Yousaf5 months ago

Offline AI + zero distractions = real deep work unlocked

Pietro Montaldo's profile picture
Pietro Montaldo5 months ago

Local models + tools are getting surprisingly powerful.

Rahul - QA - Automation - AI's profile picture
Rahul - QA - Automation - AI5 months ago

Great open source model

Scott big daddy's profile picture
Scott big daddy2 months ago

Absolutely. A little caution goes a long way. Verifying someone's identity before sending money is one of the simplest and most effective ways to protect yourself from scams.

Related Videos