Loading video...

Video Failed to Load

Go Home

IT WORKED!! Here's how you run Claude Code + Local LLM on your own machine in under 90 seconds

466,783 views • 5 months ago •via X (Twitter)

82 Comments

Agent Mo's profile picture
Agent Mo5 months ago

Just because you ran it doesn’t mean it worked. Claude code uses a lot of context tokens, you will need to find the perfect models with enough context windows

Ziwen's profile picture
Ziwen5 months ago

I showed the possibility. Now it's your turn to make your own Claude code.

Jax's profile picture
Jax5 months ago

@MoTheAgent Bold of you to assume the people that pass judgement are actually building anything let alone showing an ounce of appreciation. I will try this out and see what I can build on top , thank you.

Ziwen's profile picture
Ziwen5 months ago

@MoTheAgent Let’s go!!! Show me the sauce haha 😆

Boyuan (Nemo) Chen's profile picture
Boyuan (Nemo) Chen5 months ago

The harness is the easy part though. Where local models still fall apart is the ambiguous multi-step reasoning: 'refactor this module but don't break the tests I didn't write yet.' Would love to see someone benchmark real-world multi-file tasks, not just 90-second demos.

Ziwen's profile picture
Ziwen5 months ago

Haha maybe I should do it too

Unknown Guy 💯's profile picture
Unknown Guy 💯5 months ago

@boyuan_chen You think you can show me how it’s done ?

Ziwen's profile picture
Ziwen5 months ago

@boyuan_chen About how to set out one or

Niraj Dilshan's profile picture
Niraj Dilshan5 months ago

90 seconds to run it, 3 hours to debug the python path variables

Tim's profile picture
Tim5 months ago

Alternative with llama.cpp

LCHHAB 9's profile picture
LCHHAB 95 months ago

Congrats! Which local LLM did you pair with Claude Code?

Ziwen's profile picture
Ziwen5 months ago

I have Qwen3.5 0.8B setup. However it's because I have the Mac mini with 16GB only.

Adam Garantche's profile picture
Adam Garantche5 months ago

@LCHHAB_9 What model would you recommend if someone had more resources?

Ziwen's profile picture
Ziwen5 months ago

@LCHHAB_9 Minimax 2.7? Time will change any new model could come out

Diogenes's profile picture
Diogenes5 months ago

Slow as fuck. Unless your setup at home is super duper powerful. I mean server level memory and gpu. Else you’re wasting your time.

Ziwen's profile picture
Ziwen5 months ago

Nah I have Mac mini with 0.8 qwen 3.5 has 50 tokens / sec Pretty fast

Ziwen's profile picture
Ziwen5 months ago

16 GB by the way

Stupid As Fuck 🇧🇩's profile picture
Stupid As Fuck 🇧🇩5 months ago

Awesome!!

Ziwen's profile picture
Ziwen5 months ago

Let's go! Time to set up

Stupid As Fuck 🇧🇩's profile picture
Stupid As Fuck 🇧🇩5 months ago

I am starting in an hour so definitely setting it up!

gaurav kumar's profile picture
gaurav kumar5 months ago

Can we put azure open ai api and use it?

Ziwen's profile picture
Ziwen5 months ago

For now I don't think Leon has that setup yet.

Ziwen's profile picture
Ziwen5 months ago

Is this better? I tried to use this initially but couldn’t run it

Saikiran's profile picture
Saikiran5 months ago

Great one, but the core essence is the model than the md instructions it routes to

Ziwen's profile picture
Ziwen5 months ago

😂😂 I agree. Now we need another spy. All they need to do is putting their model in other source maps.

Paulera's profile picture
Paulera5 months ago

Claude Code is open source now so you don't need to use some shady repos

Ziwen's profile picture
Ziwen5 months ago

🤣🤣🤣 I tried with the src code it's very buggy. Asked Gemini, Codex and Minimax they all said not able to reverse engineer it...

Ziwen's profile picture
Ziwen5 months ago

I think it’s the client side code

Paulera's profile picture
Paulera5 months ago

So this is what got leaked, Claude Code client and then they open-sourced it.

Ziwen's profile picture
Ziwen5 months ago

Nah it’s been open sourced that’s why we can download with npm package

V Pathak's profile picture
V Pathak5 months ago

Just type ollama launch claude , it will do similar work as this man suggested, why everyone going crazy about leaks it's just shell not llm.

Ziwen's profile picture
Ziwen5 months ago

Oh, it was the source code itself. And someone wrapped it into a shell. So you can edit the code right away and build your own Claude code.

V Pathak's profile picture
V Pathak5 months ago

What's the purpose of building own claude code without LLM ?? And if using open source self trained model or distilled model, then only it makes sense.

Ziwen's profile picture
Ziwen5 months ago

Giving people an option. You never know what others can build with this.

V Pathak's profile picture
V Pathak5 months ago

Kudos to those individuals in advance 👏.

Ziwen's profile picture
Ziwen5 months ago

💪💪

prozic.eth's profile picture
prozic.eth5 months ago

Any guide to do it with openclaw? I don’t what my openclaw to cost me few thousands for simply prompts

Ziwen's profile picture
Ziwen5 months ago

I built one before You just need to change the config to the current model. If you want to see like a local LLM for Ollama I might come out a new video. This time can be a screen recording haha 😆

Ziwen's profile picture
Ziwen5 months ago

It doesn't work anymore cuz the repo is gone. But now check out this to implement with Claw Code

Wise's profile picture
Wise5 months ago

Nothing knew here

Ziwen's profile picture
Ziwen5 months ago

Haha, the local models are not new. But the Claude code is. We can now build our own custom Claude code.

Sebastian Buzdugan's profile picture
Sebastian Buzdugan5 months ago

cool hack, but 90 seconds setup is irrelevant if your machine overheats constantly

Ziwen's profile picture
Ziwen5 months ago

😂😂 that’s out of the concept. Haha

OneManSaas's profile picture
OneManSaas5 months ago

The 90-second setup is what matters. I've tried so many local LLM tutorials that turn into 3-hour rabbit holes with dependency hell. Actually timing yourself forces you to write instructions that work for real people, not just the person who already has everything configured.

Ziwen's profile picture
Ziwen5 months ago

💪💪 glad you think it can help.. Time to cook some sauce

Gunther Hermann's profile picture
Gunther Hermann5 months ago

thank you bing chilling

Ziwen's profile picture
Ziwen5 months ago

💪💪

0xyounes's profile picture
0xyounes5 months ago

ask him to make an app to record your screen

Ziwen's profile picture
Ziwen5 months ago

😂😂 yo i will do screen recording next time haha

__Nierowheezy's profile picture
__Nierowheezy5 months ago

We could already run this long before it was leaked

Ziwen's profile picture
Ziwen5 months ago

We can run, but we can’t build. Now we can build above on the Claude code and make your own Claude code

DissonantCat's profile picture
DissonantCat5 months ago

Or you could just use the original and change the env var ANTHROPIC_BASE_URL to point to your local setup 🤷‍♂️

Ziwen's profile picture
Ziwen5 months ago

Yea for local model! But this time we can edit the Claude Code and make your own Claude Code 😂

Osman E's profile picture
Osman E5 months ago

we need its brain not its shell

Ziwen's profile picture
Ziwen5 months ago

😂😂 who doesn’t.. we just need another spy to jump in there

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack5 months ago

@ziwenxu_ that’s dope! I’ve been wanting to try this. 90 seconds is wild, can’t wait to give it a shot!

Ziwen's profile picture
Ziwen5 months ago

Let me know if it works !!

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack5 months ago

@ziwenxu_ will do! have you tried it out yet?

Coinkong's profile picture
Coinkong5 months ago

thank you bing chilling

Ziwen's profile picture
Ziwen5 months ago

Time to cook your own Claude code haha 😆

safaa94's profile picture
safaa945 months ago

The idea is not new , it doesn't offer value ,it's just another approach for a repeated work. I wished the claude models to be free to use after the leak , but it turns out this is not possible

Ziwen's profile picture
Ziwen5 months ago

🤣🤣 Yo we need to send some spy in anthropic.

safaa94's profile picture
safaa945 months ago

Hahahaha 😄 007

Ziwen's profile picture
Ziwen5 months ago

Nothing unique just put it on the source map again 🤣🤣

P.AYREON's profile picture
P.AYREON5 months ago

Je suis en train de configurer le mien sous win 7 > llama-b8555-bin-win-cpu-x64. Et quand j'entend parler des IA ou grok qui se prends pour une iA j'en pisse de rire ..

Ziwen's profile picture
Ziwen5 months ago

Llama is bad 🤣🤣

Uncharted🦇🔊's profile picture
Uncharted🦇🔊5 months ago

@grok is this Claude code or claw?? What’s he saying exactly?? Based on how powerful Claude code is can it even be run in a single device??

2x1 Hosting's profile picture
2x1 Hosting5 months ago

BUY A PHONE STAND CHEAPO!

Ziwen's profile picture
Ziwen5 months ago

will do!!

Blaster o T's profile picture
Blaster o T5 months ago

Hell yeah! Thanks for posting!

Wen Cabrel's profile picture
Wen Cabrel5 months ago

谢谢你

Ziwen's profile picture
Ziwen5 months ago

💪💪💪 no problem

ChienSurpris's profile picture
ChienSurpris5 months ago

@grok peux tu tester et me dire si ça marche

Decoded's profile picture
Decoded5 months ago

You can already use your own llm using environmental variables on the closed source version.

Ziwen's profile picture
Ziwen5 months ago

Yup, but this time I set up with the open source version.

ClawStore's profile picture
ClawStore5 months ago

Drop full bundle so we can plug in and play!

Ziwen's profile picture
Ziwen5 months ago

😂 full bundle in the video

Vijeet Shah's profile picture
Vijeet Shah5 months ago

You can even try Echo AI also, if you want another serious coding assistant to compare. Directly download VS Code extension or npm install -g echoai

salahxyzzyx's profile picture
salahxyzzyx5 months ago

still works?

Ziwen's profile picture
Ziwen5 months ago

Yea it’s only yesterday 😂

salahxyzzyx's profile picture
salahxyzzyx5 months ago

okay am a give it a try, thanks for the info

Ziwen's profile picture
Ziwen5 months ago

💪💪

Related Videos