
Tech2Wild
@Tech2Wild • 3,377 subscribers
🎮 Tech, gaming, AI, and everything in between. 🤖 Building with it, not just talking about it. 🔥 From the mind of @ToNYD2WiLD
Shorts
Videos

This is the FASTER MODEL I've EVER Ran LOCALLY 🚀 laguna on Laguna S 2.1 · 3090 Quad INT4 • 1 parallel coding agents • Total: 3,405 tokens in 14s • Aggregate: 256.0 tok/s peak · 249.9 tok/s sustained • Per-stream: 256.0 high / 256.0 low / 256.0 avg tok/s • TTFT 0.12s avg
Tech2Wild16,404 views • 1 month ago

Running a Hermes Agent with GLM 5.2 on a Computer the size of a Deck of Playing Cards 🔥🔥🔥
Tech2Wild19,294 views • 1 month ago

2x RTX 3090s running 32 coding agents in parallel. Task: Build a trading bot (handle partial fills, rate limits, race conditions + explain reasoning). 🚀 qwen3.6-35b-a3b-autoround on 3090 35B • Total: 76,183 tokens in 64s • Aggregate: 3051.1 tok/s peak · 1183.8 tok/s sustained • Per-stream: 101.2 high / 80.5 low / 96.8 avg tok/s • TTFT 0.25s avg · E2E 55s avg
Tech2Wild15,456 views • 1 month ago

This is WILD ! 2 x 3090s 🚀 qwen3.6-35b-a3b-autoround on 3090 35B 32 parallel coding agents • Total: 25,943 tokens in 17s • Aggregate: 252275.4 tok/s peak · 1548.1 tok/s sustained • Per-stream: 20000.0 high / 1666.7 low / 9567.7 avg tok/s • TTFT 0.43s avg · E2E 17s avg
Tech2Wild14,587 views • 1 month ago
No more content to load