Loading video...
Video Failed to Load
Which models can your machine actually run? Magnitude is a 100% free, open source desktop app that: - profiles your hardware and runs sample calculations - predicts tok/s for every model before you download - recommends the best models, from fast to smart Pick your models and it handles... show more
44,616 views • 2 days ago •via X (Twitter)
29 Comments

Can't wait to fill this bad boy up

Hell yeah 🦅📈🔥

Under the hood this runs on llama.cpp today. We're replacing it with our own inference engine, built for consumer hardware: - our own kernels, optimized for each local model family - KV cache that grows and shrinks with free memory - expert offloading for models bigger than your machine Expect a release with benchmarks in the next week or two! Follow me to stay in the loop 🤝

I recently made a video on Magnitude. It's so good to see that it now comes with a desktop application. Keep shipping, guys. 🔥

The shipping will continue 🚢🔥

Cool.

Thank you sir

Is it possible to move to this version by updating from a previous npm i -g magnitude install?

If you download the new app it will automatically set itself up and disconnect the old CLI!

I'd have to prepare an Arch Linux package to install it cleanly. Are you planing on maintaing an AUR package for magnitude?

If you can PR one 😄 this is something I want to support

I am not promising anything, but I might take a look at how other similar PKGBUILDS are created and compare to your tar.gz Linux package.

Amazing that’s what I was looking for !

Let me know how it goes!

Will give it a try on my M4 MacMini after LM Studio gets more unstable with their recent updates.

Does it work on Windows within WSL? But for VS Code and OpenClaw Desktop on native Windows, all my developer tooling lives in WSL Ubuntu.

Yes, it works on WSL! If your agents run there too, they connect directly. If they run on Windows, they’d connect through WSL’s localhost forwarding. Are the agents themselves running in WSL, or just accessing your projects there?

Does it support hybrid quantized models?

Yes we support mixed-precision GGUF (like Unsloth Dynamic quants). Any particular model you have in mind?

current top 3 production local models right now are : 1-glm 5.3 flash 2-Deepseek v4 & v4 vision and v4.1 .3-Qwen 3.8 flash and waiting for new qwen 3.8 omni wights release 4-qwen 3.8 27b but there are lots of hybrid quantazaion and ablitraitons"

Kinda sad you guys did away with the CLI I really liked it!

The desktop app is bundled with the headless CLI so your agent can still drive it! But yes we did away with the TUI :( desktop app is just much more expressive UI

Ohh

Looks like you can specify were the model directory can be on the workstation using Magnitude?

Not currently but we have an issue open for it - I'm going to implement this today

The “predict before you download” step is underrated—local model experimentation gets much less frustrating when hardware fit is explicit up front. Open source also makes the recommendations auditable instead of magic.
55 models in the whole catalog? Seems like this exists to gather and exfiltrate your LLM model data to the company so they can build their model database up.

Can you incorporate free token attributes to run models in vram and ram?

an honest recommendation engine for loocal models thank you

