Alexey Fateev's banner
Alexey Fateev's profile picture

Alexey Fateev

@superalesha1,840 subscribers

⚡I benchmark local LLMs on 4x RTX 3090s. exact configs, tok/s, VRAM, and what broke. ❤️ https://t.co/tIPzthsdkv - 2xDGX Spark 🚀96GB VRAM | Local AI

Shorts

I think I need professional medical help. I cant stop. Told myself just one more and that was 40 videos ago. Until Qwen3.8 27B drops my 4x3090 will keep cooking these until they die. More videos and the prompts are in the replies.

I think I need professional medical help. I cant stop. Told myself just one more and that was 40 videos ago. Until Qwen3.8 27B drops my 4x3090 will keep cooking these until they die. More videos and the prompts are in the replies.

83,692 次观看

Minimax H3, same 4x3090 rig, same clip. Yesterday it rendered in 11:21. Today it renders in 3:45 🤯 The hardware did not change. Two things happened overnight. An AI agent rewrote the attention CUDA kernel and hit a half speed trap inside GeForce tensor cores that most people never heard of. Then a stranger on HuggingFace dropped a LoRA that cuts 20 sampling steps down to 4. One of these two mattered way more than the other. Breakdown below. I packed all of it into ready ComfyUI workflows for my rig. If you want them, ask in the replies and I'll share.

Minimax H3, same 4x3090 rig, same clip. Yesterday it rendered in 11:21. Today it renders in 3:45 🤯 The hardware did not change. Two things happened overnight. An AI agent rewrote the attention CUDA kernel and hit a half speed trap inside GeForce tensor cores that most people never heard of. Then a stranger on HuggingFace dropped a LoRA that cuts 20 sampling steps down to 4. One of these two mattered way more than the other. Breakdown below. I packed all of it into ready ComfyUI workflows for my rig. If you want them, ask in the replies and I'll share.

68,011 次观看

Videos

没有更多内容可加载