正在加载视频...
视频加载失败
Can you spot the difference? 👀 Speculative decoding helps accelerate generative AI at the edge, especially on NVIDIA Jetson. Qwen 3.8 27B went from 13 to 35 tokens/sec, while Nemotron 3.5 Lightning went from 65 to 115. See how it works:
44,587 次观看 • 6 天前 •via X (Twitter)
0 条评论
暂无评论
原始帖子的评论将显示在这里
相关视频
0:30
Sensitive content
Qwen 3.8 Max beat Fable 5 at building 3D physics scenes for 7x cheaper! We gave two models the same task. Build three self-contained 3D scenes, each one HTML file with real physics that runs itself. Prompts: - A marble machine that lifts marbles up a wheel and drops them on a loop - A car factory assembly line - A sawmill cutting logs into planks Outputs: Qwen 3.8 Max: 46.6K tokens, $0.28 Fable 5: 38.7K tokens, $1.93 Qwen worked much harder on the details. Its textures and shadows look real. Fable 5 built a fast prototype that works and left it there. In the marble machine, Fable's marbles show up from nowhere at the top of the wheel. In Qwen's, you can see them go up. On the car line, Qwen's cars look much better and more real. Fable's sawmill is very strange. The saw looks wrong, and the log goes right through a beam before the saw cuts it. Atomic Chat will have day zero support to run Qwen 3.8 Max locally!
atomic.chat
210,569 次观看 • 22 天前

