Loading video...
Video Failed to Load
Wait Xiaomi has also released the best 9B model in the world?! This model is based on Qwen-9B... And you can easily run it even on 8GB of memory! - MIT license - Multimodal (image + text) - Solid for coding/agentic tasks Perfect to use locally and offline... even... show more
82,735 views • 1 day ago •via X (Twitter)
24 Comments

Base model weights: Quantized GGUF:

Stop making wrong claims and that 8GB RAM thing is just so stupid about you

Fits using a Q4 version, especially if you're offloading the context in the ram

Amazing. Now let's see your tests on 8gb VRAM 🙃

我认为并不是,Ornith-1.5-9B在编码任务上明显更强

@grok will this work on my pc? I have rtx 3060 and 32 GB RAM

max-distilled fable.... nice.

Everyone is distilling from everyone... and even closed labs are integrating open research.

yes. IMO anthropic has been distilled most. Meta distills least from other models. A bit like inbreeding. We need more diversity!

This one is very good and high quality!

The 8GB angle is what makes this genuinely useful, not just impressive on a benchmark. I am curious how it holds up on longer agentic tasks once context and tool calls start competing for memory.

Eight-gigabyte multimodal inference makes privacy concrete, but leaves little context headroom.

@grok how would this compare to spark 2.5 4b in terms of performance and requirements? If I have 12bg if VRAM, what tokens a second could I expect?

Have you tested it?

On it. So far so good.

That's really impressive. Open-source, low-resource requirement, multi-modal, great for local deployment.

can it actually hold a long agentic loop or does it fall apart after step three

GPUs go burrr is optional now. A 9B MIT-licensed agentic model that runs offline on 8GB of unified memory is the actual news here.

I think bonsai 2 will still be on top due to the 27B representational space .

Will it work on non Mac

unfortunately the RL version is not yet released :-) only the SFT versino

how does it handle complex coding tasks compared to other models?

Wait... Fuck off

is it actually any good though

