4 DGX Sparks running DeepSeek V4.1 Flash 6 concurrent...

Wësche's profile picture

Wësche

23,510 次观看 • 5 天前

Okay, I guess I didn't waste all that money...

Killy's profile picture

Killy

43,181 次观看 • 5 天前

Day-0 Qwen3.8-27B vs Qwen3.6-27B: the voxel pagoda test. Same...

Wësche's profile picture

Wësche

97,988 次观看 • 1 个月前

One task, two models. GLM 5.3 Flash running slower...

Mia's profile picture

Mia

27,249 次观看 • 9 天前

vllm-exl3 v0.3.0 is LIVE with custom native CUDA kernels...

Cruz's profile picture

Cruz

21,173 次观看 • 12 天前

nvidia/Qwen3.6-35B-A3B-NVFP4 running in vLLM nightly on my Nvidia GB10...

Onur Solmaz's profile picture

Onur Solmaz

27,887 次观看 • 3 个月前

Two models. One DGX Spark One identical voxel Eiffel...

Wësche's profile picture

Wësche

51,404 次观看 • 27 天前

Qwen3.8-Flash-Next now reaches ~43 tok/s after a 122,902-token prompt...

Cruz's profile picture

Cruz

12,258 次观看 • 8 天前

I think the CMP 170HX just became one of...

Own Your Compute's profile picture

Own Your Compute

18,858 次观看 • 23 天前

This is awesome. Thanks Mia for this Was able...

Melvin Vivas's profile picture

Melvin Vivas

99,403 次观看 • 15 天前

Ornith-1.0 35B - MLX Device - 128gb m5 max...

Trevor Wood's profile picture

Trevor Wood

11,765 次观看 • 1 个月前

This is the worst model I have ever tested,...

Wësche's profile picture

Wësche

18,306 次观看 • 21 天前

1,000 tok/s vs 85 tok/s visualized

nader dabit's profile picture

nader dabit

211,889 次观看 • 2 个月前

Three Cline agents. Same 120B model. One rule: terminate...

Cline's profile picture

Cline

14,782 次观看 • 5 个月前

167 tok/s on a single RTX 4090. FreeToken just...

FHILY👑's profile picture

FHILY👑

33,058 次观看 • 21 天前

Laguna XS 2.1 just matched Qwen 3.6 35B on...

0xMarioNawfal's profile picture

0xMarioNawfal

63,839 次观看 • 1 个月前

Qwen3.8-Flash-Next is starting to feel like the local model...

FHILY👑's profile picture

FHILY👑

20,253 次观看 • 19 天前

Deepseek v4.1 Flash on 24GB Macbook. I think I...

Marco Franzon's profile picture

Marco Franzon

47,585 次观看 • 2 天前

NVIDIA just dropped Nemotron-3-Nano:4b — a tiny 2.8GB model....

stevibe's profile picture

stevibe

127,570 次观看 • 6 个月前

If you have an RTX 3090 or 4090, Mia...

Yume_X's profile picture

Yume_X

38,511 次观看 • 12 天前

🎉 Congrats to Thinking Machines on TML Inkling—a 1T-parameter...

vLLM's profile picture

vLLM

49,464 次观看 • 2 个月前