Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

IT WORKED!! Here's how you run Claude Code + Local LLM on your own machine in under 90 seconds

466,783 Aufrufe • vor 5 Monaten •via X (Twitter)

82 Kommentare

Profilbild von Agent Mo
Agent Movor 5 Monaten

Just because you ran it doesn’t mean it worked. Claude code uses a lot of context tokens, you will need to find the perfect models with enough context windows

Profilbild von Ziwen
Ziwenvor 5 Monaten

I showed the possibility. Now it's your turn to make your own Claude code.

Profilbild von Jax
Jaxvor 5 Monaten

@MoTheAgent Bold of you to assume the people that pass judgement are actually building anything let alone showing an ounce of appreciation. I will try this out and see what I can build on top , thank you.

Profilbild von Ziwen
Ziwenvor 5 Monaten

@MoTheAgent Let’s go!!! Show me the sauce haha 😆

Profilbild von Boyuan (Nemo) Chen
Boyuan (Nemo) Chenvor 5 Monaten

The harness is the easy part though. Where local models still fall apart is the ambiguous multi-step reasoning: 'refactor this module but don't break the tests I didn't write yet.' Would love to see someone benchmark real-world multi-file tasks, not just 90-second demos.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Haha maybe I should do it too

Profilbild von Unknown Guy 💯
Unknown Guy 💯vor 5 Monaten

@boyuan_chen You think you can show me how it’s done ?

Profilbild von Ziwen
Ziwenvor 5 Monaten

@boyuan_chen About how to set out one or

Profilbild von Niraj Dilshan
Niraj Dilshanvor 5 Monaten

90 seconds to run it, 3 hours to debug the python path variables

Profilbild von Tim
Timvor 5 Monaten

Alternative with llama.cpp

Profilbild von LCHHAB 9
LCHHAB 9vor 5 Monaten

Congrats! Which local LLM did you pair with Claude Code?

Profilbild von Ziwen
Ziwenvor 5 Monaten

I have Qwen3.5 0.8B setup. However it's because I have the Mac mini with 16GB only.

Profilbild von Adam Garantche
Adam Garantchevor 5 Monaten

@LCHHAB_9 What model would you recommend if someone had more resources?

Profilbild von Ziwen
Ziwenvor 5 Monaten

@LCHHAB_9 Minimax 2.7? Time will change any new model could come out

Profilbild von Diogenes
Diogenesvor 5 Monaten

Slow as fuck. Unless your setup at home is super duper powerful. I mean server level memory and gpu. Else you’re wasting your time.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Nah I have Mac mini with 0.8 qwen 3.5 has 50 tokens / sec Pretty fast

Profilbild von Ziwen
Ziwenvor 5 Monaten

16 GB by the way

Profilbild von Stupid As Fuck 🇧🇩
Stupid As Fuck 🇧🇩vor 5 Monaten

Awesome!!

Profilbild von Ziwen
Ziwenvor 5 Monaten

Let's go! Time to set up

Profilbild von Stupid As Fuck 🇧🇩
Stupid As Fuck 🇧🇩vor 5 Monaten

I am starting in an hour so definitely setting it up!

Profilbild von gaurav kumar
gaurav kumarvor 5 Monaten

Can we put azure open ai api and use it?

Profilbild von Ziwen
Ziwenvor 5 Monaten

For now I don't think Leon has that setup yet.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Is this better? I tried to use this initially but couldn’t run it

Profilbild von Saikiran
Saikiranvor 5 Monaten

Great one, but the core essence is the model than the md instructions it routes to

Profilbild von Ziwen
Ziwenvor 5 Monaten

😂😂 I agree. Now we need another spy. All they need to do is putting their model in other source maps.

Profilbild von Paulera
Pauleravor 5 Monaten

Claude Code is open source now so you don't need to use some shady repos

Profilbild von Ziwen
Ziwenvor 5 Monaten

🤣🤣🤣 I tried with the src code it's very buggy. Asked Gemini, Codex and Minimax they all said not able to reverse engineer it...

Profilbild von Ziwen
Ziwenvor 5 Monaten

I think it’s the client side code

Profilbild von Paulera
Pauleravor 5 Monaten

So this is what got leaked, Claude Code client and then they open-sourced it.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Nah it’s been open sourced that’s why we can download with npm package

Profilbild von V Pathak
V Pathakvor 5 Monaten

Just type ollama launch claude , it will do similar work as this man suggested, why everyone going crazy about leaks it's just shell not llm.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Oh, it was the source code itself. And someone wrapped it into a shell. So you can edit the code right away and build your own Claude code.

Profilbild von V Pathak
V Pathakvor 5 Monaten

What's the purpose of building own claude code without LLM ?? And if using open source self trained model or distilled model, then only it makes sense.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Giving people an option. You never know what others can build with this.

Profilbild von V Pathak
V Pathakvor 5 Monaten

Kudos to those individuals in advance 👏.

Profilbild von Ziwen
Ziwenvor 5 Monaten

💪💪

Profilbild von prozic.eth
prozic.ethvor 5 Monaten

Any guide to do it with openclaw? I don’t what my openclaw to cost me few thousands for simply prompts

Profilbild von Ziwen
Ziwenvor 5 Monaten

I built one before You just need to change the config to the current model. If you want to see like a local LLM for Ollama I might come out a new video. This time can be a screen recording haha 😆

Profilbild von Ziwen
Ziwenvor 5 Monaten

It doesn't work anymore cuz the repo is gone. But now check out this to implement with Claw Code

Profilbild von Wise
Wisevor 5 Monaten

Nothing knew here

Profilbild von Ziwen
Ziwenvor 5 Monaten

Haha, the local models are not new. But the Claude code is. We can now build our own custom Claude code.

Profilbild von Sebastian Buzdugan
Sebastian Buzduganvor 5 Monaten

cool hack, but 90 seconds setup is irrelevant if your machine overheats constantly

Profilbild von Ziwen
Ziwenvor 5 Monaten

😂😂 that’s out of the concept. Haha

Profilbild von OneManSaas
OneManSaasvor 5 Monaten

The 90-second setup is what matters. I've tried so many local LLM tutorials that turn into 3-hour rabbit holes with dependency hell. Actually timing yourself forces you to write instructions that work for real people, not just the person who already has everything configured.

Profilbild von Ziwen
Ziwenvor 5 Monaten

💪💪 glad you think it can help.. Time to cook some sauce

Profilbild von Gunther Hermann
Gunther Hermannvor 5 Monaten

thank you bing chilling

Profilbild von Ziwen
Ziwenvor 5 Monaten

💪💪

Profilbild von 0xyounes
0xyounesvor 5 Monaten

ask him to make an app to record your screen

Profilbild von Ziwen
Ziwenvor 5 Monaten

😂😂 yo i will do screen recording next time haha

Profilbild von __Nierowheezy
__Nierowheezyvor 5 Monaten

We could already run this long before it was leaked

Profilbild von Ziwen
Ziwenvor 5 Monaten

We can run, but we can’t build. Now we can build above on the Claude code and make your own Claude code

Profilbild von DissonantCat
DissonantCatvor 5 Monaten

Or you could just use the original and change the env var ANTHROPIC_BASE_URL to point to your local setup 🤷‍♂️

Profilbild von Ziwen
Ziwenvor 5 Monaten

Yea for local model! But this time we can edit the Claude Code and make your own Claude Code 😂

Profilbild von Osman E
Osman Evor 5 Monaten

we need its brain not its shell

Profilbild von Ziwen
Ziwenvor 5 Monaten

😂😂 who doesn’t.. we just need another spy to jump in there

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 5 Monaten

@ziwenxu_ that’s dope! I’ve been wanting to try this. 90 seconds is wild, can’t wait to give it a shot!

Profilbild von Ziwen
Ziwenvor 5 Monaten

Let me know if it works !!

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 5 Monaten

@ziwenxu_ will do! have you tried it out yet?

Profilbild von Coinkong
Coinkongvor 5 Monaten

thank you bing chilling

Profilbild von Ziwen
Ziwenvor 5 Monaten

Time to cook your own Claude code haha 😆

Profilbild von safaa94
safaa94vor 5 Monaten

The idea is not new , it doesn't offer value ,it's just another approach for a repeated work. I wished the claude models to be free to use after the leak , but it turns out this is not possible

Profilbild von Ziwen
Ziwenvor 5 Monaten

🤣🤣 Yo we need to send some spy in anthropic.

Profilbild von safaa94
safaa94vor 5 Monaten

Hahahaha 😄 007

Profilbild von Ziwen
Ziwenvor 5 Monaten

Nothing unique just put it on the source map again 🤣🤣

Profilbild von P.AYREON
P.AYREONvor 5 Monaten

Je suis en train de configurer le mien sous win 7 > llama-b8555-bin-win-cpu-x64. Et quand j'entend parler des IA ou grok qui se prends pour une iA j'en pisse de rire ..

Profilbild von Ziwen
Ziwenvor 5 Monaten

Llama is bad 🤣🤣

Profilbild von Uncharted🦇🔊
Uncharted🦇🔊vor 5 Monaten

@grok is this Claude code or claw?? What’s he saying exactly?? Based on how powerful Claude code is can it even be run in a single device??

Profilbild von 2x1 Hosting
2x1 Hostingvor 5 Monaten

BUY A PHONE STAND CHEAPO!

Profilbild von Ziwen
Ziwenvor 5 Monaten

will do!!

Profilbild von Blaster o T
Blaster o Tvor 5 Monaten

Hell yeah! Thanks for posting!

Profilbild von Wen Cabrel
Wen Cabrelvor 5 Monaten

谢谢你

Profilbild von Ziwen
Ziwenvor 5 Monaten

💪💪💪 no problem

Profilbild von ChienSurpris
ChienSurprisvor 5 Monaten

@grok peux tu tester et me dire si ça marche

Profilbild von Decoded
Decodedvor 5 Monaten

You can already use your own llm using environmental variables on the closed source version.

Profilbild von Ziwen
Ziwenvor 5 Monaten

Yup, but this time I set up with the open source version.

Profilbild von ClawStore
ClawStorevor 5 Monaten

Drop full bundle so we can plug in and play!

Profilbild von Ziwen
Ziwenvor 5 Monaten

😂 full bundle in the video

Profilbild von Vijeet Shah
Vijeet Shahvor 5 Monaten

You can even try Echo AI also, if you want another serious coding assistant to compare. Directly download VS Code extension or npm install -g echoai

Profilbild von salahxyzzyx
salahxyzzyxvor 5 Monaten

still works?

Profilbild von Ziwen
Ziwenvor 5 Monaten

Yea it’s only yesterday 😂

Profilbild von salahxyzzyx
salahxyzzyxvor 5 Monaten

okay am a give it a try, thanks for the info

Profilbild von Ziwen
Ziwenvor 5 Monaten

💪💪

Ähnliche Videos