Loading video...

Video Failed to Load

Go Home

Hot tip for anyone doing AI dev: Use Ollama to easily run models like Deepseek-r1 or Gemma locally on your machine. It downloads them and spins up a server with an OpenAI SDK compatible API The smaller models are fast and good enough to work on new features or...

153,435 views • 1 year ago •via X (Twitter)

11 Comments

Tanner Linsley's profile picture
Tanner Linsley1 year ago

You're so cool. I wish I could be this level of cool.

Wes Bos's profile picture
Wes Bos1 year ago

I only hope to be the level of cool that it takes to have a stack named after me

Rahim Nathwani's profile picture
Rahim Nathwani1 year ago

If you're doing `ollama run deepseek-r1` you're NOT running a Deepseek R1 model, but a Qwen3 8B param model finetuned on outputs from R1. Deepseek R1 has 671B params, requiring >700GB of (V)RAM. There's a 1.58-bit quantized version that fits in 131GB.

Dennis's profile picture
Dennis1 year ago

@lmstudio is another great option.

goosewin's profile picture
goosewin1 year ago

In my experience, these models are rarely good enough for solving the kinds of problems I face day to day

🐧 lalo adrian morales 𝕏's profile picture
🐧 lalo adrian morales 𝕏1 year ago

been @ollama #1 fan since day 1!

Arvid Kahl's profile picture
Arvid Kahl1 year ago

This kind of AI-related work will be why many devs will go back to desktop computers with solid GPUs. The speed you get our of a well-quantized model like the one you use w/ Ollama on a high-end Mac Studio is magical. That's what I'm using, and it rocks :D

Peter Cruckshank's profile picture
Peter Cruckshank1 year ago

I feel like I don't think I've got enough VRAM to run a good enough local model 😮‍💨 Guess that means UPGRADE!! 🎉🙌🏻 This does look like it would be fun to try out on smaller features.

рома's profile picture
рома1 year ago

What do you think is the best small model to run locally? I tried only Mistral yet

zuri-zel's profile picture
zuri-zel1 year ago

what about using lmstudio? have you tried both?

Wes Bos's profile picture
Wes Bos1 year ago

the app, or the SDK? I haven't tried either. Any good?

Related Videos

how you can use openAI codex & gpt 5.5 completely FREE (the full guide) 100% legit. no subscription, zero API cost. up to 1M+ token/day. you need just an openAI account and here's how to set it up in 5mins. openAI has a program that gives eligible developers free API usage every day in exchange for sharing API data that helps improve future models. it's not a one-time credit, your allowance refreshes daily. depending on your usage tier, you can get access to hundreds of thousands, or even millions, of free tokens every single day on supported models. here's how to activate it: 1️⃣open your API dashboard: 2️⃣go to settings → data controls 3️⃣enable data sharing for your organization or project 4️⃣make sure your account has a positive API balance 5️⃣save the settings if your account is eligible, you'll see a message confirming access to complimentary daily usage. before you turn it on, know the tradeoff: • prompts and outputs from shared projects can be used to improve openai's models • don't use it for confidential information, client work, or sensitive data • eligibility depends on your account type and settings for everyone else, it's an incredible deal. use it to: • learn AI development • build side projects • experiment with codex • test agents and automations • prototype ideas without worrying about API costs most developers burn money testing ideas. this lets you experiment at scale while spending little to nothing.

m0h

68,972 views • 1 month ago