Video wird geladen...
Video konnte nicht geladen werden
hy4 preview changes the question from “can a lab run 770B?” to “can your team stand it up?” tencent hunyuan released an open w8 text model: 𝟕𝟕𝟎𝐁 𝐭𝐨𝐭𝐚𝐥, 𝟒𝟗𝐁 𝐚𝐜𝐭𝐢𝐯𝐞 𝐩𝐞𝐫 𝐭𝐨𝐤𝐞𝐧, Apache 2.0 and native 𝟏𝐦 𝐜𝐨𝐧𝐭𝐞𝐱𝐭. Day 0 paths exist for vLLM and SGLang, plus an official... show more
32,855 Aufrufe • vor 24 Tagen •via X (Twitter)
11 Kommentare

Liam | AI Tools & Newsvor 23 Tagen
Amazing

Maxvor 24 Tagen
Really enjoyed this one. There’s a thoughtful perspective here that genuinely stayed with me.

NOVAvor 24 Tagen
214GiB mixed quant is the part that actually matters

Iris Hayesvor 23 Tagen
Apache 2.0 plus 1M context is a serious ownership move

Liamvor 23 Tagen
Patched llama.cpp will slow early testers

Amber Nexusvor 23 Tagen
Masterpiece

Vikas guptavor 23 Tagen
Need real coding and long-context numbers, not just size claims

Rachel Woodsvor 24 Tagen
Still not a laptop model no matter how you slice 214GiB

Brian Haduvor 24 Tagen
if your team can't handle it, deploying 770B won't matter at all

Muhammad Alivor 24 Tagen
49B active is what makes this class even serveable

Shahid Ansarivor 24 Tagen
Tokens per second on your own box is the only score that counts.
