4 DGX Sparks running DeepSeek V4.1 Flash 6 concurrent...

Wësche's profile picture

Wësche

23,510 просмотров • 5 дней назад

Okay, I guess I didn't waste all that money...

Killy's profile picture

Killy

43,181 просмотров • 5 дней назад

Day-0 Qwen3.8-27B vs Qwen3.6-27B: the voxel pagoda test. Same...

Wësche's profile picture

Wësche

97,988 просмотров • 1 месяц назад

One task, two models. GLM 5.3 Flash running slower...

Mia's profile picture

Mia

27,249 просмотров • 9 дней назад

vllm-exl3 v0.3.0 is LIVE with custom native CUDA kernels...

Cruz's profile picture

Cruz

21,173 просмотров • 12 дней назад

nvidia/Qwen3.6-35B-A3B-NVFP4 running in vLLM nightly on my Nvidia GB10...

Onur Solmaz's profile picture

Onur Solmaz

27,887 просмотров • 3 месяцев назад

Two models. One DGX Spark One identical voxel Eiffel...

Wësche's profile picture

Wësche

51,404 просмотров • 27 дней назад

Qwen3.8-Flash-Next now reaches ~43 tok/s after a 122,902-token prompt...

Cruz's profile picture

Cruz

12,258 просмотров • 8 дней назад

I think the CMP 170HX just became one of...

Own Your Compute's profile picture

Own Your Compute

18,858 просмотров • 23 дней назад

This is awesome. Thanks Mia for this Was able...

Melvin Vivas's profile picture

Melvin Vivas

99,403 просмотров • 15 дней назад

Ornith-1.0 35B - MLX Device - 128gb m5 max...

Trevor Wood's profile picture

Trevor Wood

11,765 просмотров • 1 месяц назад

This is the worst model I have ever tested,...

Wësche's profile picture

Wësche

18,306 просмотров • 21 дней назад

1,000 tok/s vs 85 tok/s visualized

nader dabit's profile picture

nader dabit

211,889 просмотров • 2 месяцев назад

Three Cline agents. Same 120B model. One rule: terminate...

Cline's profile picture

Cline

14,782 просмотров • 5 месяцев назад

167 tok/s on a single RTX 4090. FreeToken just...

FHILY👑's profile picture

FHILY👑

33,058 просмотров • 21 дней назад

Laguna XS 2.1 just matched Qwen 3.6 35B on...

0xMarioNawfal's profile picture

0xMarioNawfal

63,839 просмотров • 1 месяц назад

Qwen3.8-Flash-Next is starting to feel like the local model...

FHILY👑's profile picture

FHILY👑

20,253 просмотров • 19 дней назад

Deepseek v4.1 Flash on 24GB Macbook. I think I...

Marco Franzon's profile picture

Marco Franzon

47,585 просмотров • 2 дней назад

NVIDIA just dropped Nemotron-3-Nano:4b — a tiny 2.8GB model....

stevibe's profile picture

stevibe

127,570 просмотров • 6 месяцев назад

If you have an RTX 3090 or 4090, Mia...

Yume_X's profile picture

Yume_X

38,511 просмотров • 12 дней назад

🎉 Congrats to Thinking Machines on TML Inkling—a 1T-parameter...

vLLM's profile picture

vLLM

49,464 просмотров • 2 месяцев назад