Muse Glimmer, A 30B parameter dense model swallowing a...

Alok's profile picture

Alok

65,480 views • 1 month ago

Qwen 3.8 27B (dense) running on a single RTX...

Alok's profile picture

Alok

383,369 views • 1 month ago

50% more context unlocked for Qwen 3.8 27b Q4_K_XL...

Alok's profile picture

Alok

39,189 views • 26 days ago

Qwen 3.8 27B Q4_K_M - 90 tokens/sec on a...

Alok's profile picture

Alok

105,388 views • 28 days ago

If you thought the Gemma 4 31B (dense) model...

Alok's profile picture

Alok

40,993 views • 1 month ago

The "I don't have enough VRAM" excuse just died....

Alok's profile picture

Alok

19,370 views • 1 month ago

Gemma 4 26B A4B MoE - 500+ t/s decode...

Alok's profile picture

Alok

17,465 views • 1 month ago

Gemma 4 12B QAT (dense) achieves 1000+ tokens/sec prefill...

Alok's profile picture

Alok

34,500 views • 3 months ago

Open source AI is actually moving at an unhinged...

Alok's profile picture

Alok

26,841 views • 3 months ago

the 24gb vram tier is enough for most builder...

Sudo su's profile picture

Sudo su

19,576 views • 4 months ago

I told you to claim your free 16GB NVIDIA...

Alok's profile picture

Alok

170,442 views • 2 months ago

you're paying $20/mo for something your $500 GPU can...

Alok's profile picture

Alok

36,691 views • 2 months ago

Run Gemma 4 26b MTP on 8 GB VRAM...

Alok's profile picture

Alok

200,913 views • 3 months ago

Run Gemma 4 26B MoE on 8GB VRAM with...

Alok's profile picture

Alok

292,770 views • 3 months ago

my 8 GB VRAM gaming laptop is absolutely going...

Alok's profile picture

Alok

63,689 views • 3 months ago

90% of "AI developers" just download pre packaged GGUF...

Alok's profile picture

Alok

62,631 views • 2 months ago

If you have an RTX 3090 or 4090, Mia...

Yume_X's profile picture

Yume_X

38,685 views • 13 days ago

Qwen3.8-Flash-Next now reaches ~43 tok/s after a 122,902-token prompt...

Cruz's profile picture

Cruz

12,258 views • 8 days ago

I just got Gemma 4 26B A4B MoE model...

Alok's profile picture

Alok

105,094 views • 2 months ago

Free NVIDIA GPU with 16 GB VRAM GPU for...

Alok's profile picture

Alok

182,483 views • 2 months ago

llama.cpp isn't just for text LLMs anymore. Pure C++...

Alok's profile picture

Alok

47,881 views • 1 month ago