正在加载视频...

视频加载失败

Unplugging completely! No WiFi and zero notifications. A great way to get deep focus on a project. Here is a walkthrough showing how to run Gemma 4 (26B A4B) fully offline with LM Studio & OpenCode to parse PDFs, ask questions, and build sites 100% locally.

236,747 次观看 • 5 个月前 •via X (Twitter)

31 条评论

Google Gemma 的头像
Google Gemma5 个月前

Want to take your coding setup anywhere? This video, with @ianballantyne from the Gemma team, shows how to code off the grid by serving Gemma 4 on LM Studio/Ollama, connecting OpenCode via an OpenAI endpoint, using liteparse for HTML/CSS/JS, and managing context memory offline.

Raditedu 的头像
Raditedu5 个月前

Had a emergency situation on the farm last week, was in an area with no cellphone signal. Urgently needed to find out what a term in documentation of animal medication ment. Realised I tested Gemma 4 on my phone, asked Gemma and got an IMMEDIATE offline answer of what the term ment, saved animal’s life 🙏

KellyV 的头像
KellyV5 个月前

@lmstudio On your IPhone AI Agent by Gemma4 100% Open Source 100% Runs locally with no internet required, protecting privacy

Iqbal Bhawana 的头像
Iqbal Bhawana5 个月前

The fact of the matter is that no 35 billion, 27 billion, or 26 billion model—whether quantized to Q4 or Q3 for that matter—is currently deployable on MacBooks, even if these MacBooks have compatible memory as some people are claiming. I have a MacBook M3 Pro with 36 GB RAM and I have not been able to run these models, even the Q4 GGUF models or MLX models for that matter (be it Gemma 4, Qwen3.6, or whatever). You need to have at least 64 GB RAM to get a decent output from these models. That is the bottom line.

Konstantin Gladych 的头像
Konstantin Gladych5 个月前

Nice case! would like to discover local agents with Gemma4 hosted by @atomic_chat_hq they have TurboQuant implemented and MLX as well. Perfect for large context window tasks.

Steve Oak 的头像
Steve Oak5 个月前

counterpoint: gemma 4 running locally is cool but i know for a fact i would still find a way to get distracted by a pre-installed calculator app or something 💀

MyNextDeveloper 的头像
MyNextDeveloper5 个月前

The fact that a 26B model runs fully locally now tells us everything about how fast this space is moving.

Hesu Krypto 的头像
Hesu Krypto5 个月前

What's the system requirement for this?

Sandipan Kundu 🛠 的头像
Sandipan Kundu 🛠5 个月前

Or, you can just use Friday.

deko 的头像
deko5 个月前

Nice you advertise opencode here but ban people if they use it with gemini afaik

Xing Han Lu 的头像
Xing Han Lu5 个月前

Which one do you use in-house when using gemma as an agent? Opencode, Pi or Openhands?

Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽 的头像
Avatar of Khaine 🇪🇺🇺🇦🇹🇼🇬🇪🗽5 个月前

From my experience Gemma4 is a maybe OK as a chat interface, but a mess as coding agent and for tool calling. Messes up tags all the time an gets hung up in infinite loops frequently. I had to revert to qwen 3.5.

Julius 的头像
Julius5 个月前

I’ve noticed going fully offline really sharpens focus, this kind of setup makes that actually practical

Trevor Sullivan 的头像
Trevor Sullivan5 个月前

All you need is a laptop, a solar panel, and a USB-C cable. 🗻 #MountainCoder #MountainDev #MaybeStarlinkToo

alfredoSauce 的头像
alfredoSauce5 个月前

Enjoying it on iOS so far.

Coffee fans 的头像
Coffee fans5 个月前

Local 26B still chokes on RAM. Quantize to 4-bit or disk swapping will wreck your throughput.

Samian 的头像
Samian5 个月前

unplugged local runs hit different. ollama vibes all day

fajar sp 的头像
fajar sp5 个月前

if i got 16gb ram, which model is perfect for me ?

DuoNeural 的头像
DuoNeural5 个月前

🤯The Gemma 4 models are next level. The intelligence packed in is insane for the parameter counts. I think my favorite is Gemma 4 E4B because its small and scrappy. I can fit two of them in 8GB VRAM with 131k context windows using the versions we quantized to 4 bit at 30+TPS at: What's everyone elses favorite flavor of Gemma 4? Pls respond in comments 👇

Vladimir Tchuiev 的头像
Vladimir Tchuiev5 个月前

So you've given up completely on Gemini CLI?

Daniel Schemann 的头像
Daniel Schemann5 个月前

Gemma 4 is a game-changer! Before, I was required to use cloud providers, now I am able to directly integrate it into critical parts of my app that now runs locally, doing tasks that bigger models did before.

DZ 的头像
DZ5 个月前

@lmstudio 如果你顺便提一下流畅运行需要什么样的硬件配置,我想大家的兴趣就没那么大了. 至少要25tps级别

Far 的头像
Far5 个月前

finally a tutorial that doesn't require a cloud subscription to break your wifi router

Zchat | Shielded messenger 的头像
Zchat | Shielded messenger5 个月前

offline inference is the only architecture where privacy is enforced by the hardware, not promised by the terms of service

beartcoiner 🟢🟢🔴🔴 的头像
beartcoiner 🟢🟢🔴🔴5 个月前

Where can I get a no-Woke AI?

Aaron (Youshen) Lim 的头像
Aaron (Youshen) Lim5 个月前

This is very interesting demo of local, edge deployment using the larger Gemma 4 26B A4B. Once I get better hardware & memory (currently 16GB), I would love to give this a try.

Sid Sanyal 的头像
Sid Sanyal5 个月前

🔥👌🤩

Arslan Yousaf 的头像
Arslan Yousaf5 个月前

Offline AI + zero distractions = real deep work unlocked

Pietro Montaldo 的头像
Pietro Montaldo5 个月前

Local models + tools are getting surprisingly powerful.

Rahul - QA - Automation - AI 的头像
Rahul - QA - Automation - AI5 个月前

Great open source model

Scott big daddy 的头像
Scott big daddy2 个月前

Absolutely. A little caution goes a long way. Verifying someone's identity before sending money is one of the simplest and most effective ways to protect yourself from scams.

相关视频