Загрузка видео...
Не удалось загрузить видео
Can you spot the difference? 👀 Speculative decoding helps accelerate generative AI at the edge, especially on NVIDIA Jetson. Qwen 3.8 27B went from 13 to 35 tokens/sec, while Nemotron 3.5 Lightning went from 65 to 115. See how it works:
44,587 просмотров • 6 дней назад •via X (Twitter)
Комментарии: 0
Нет доступных комментариев
Здесь появятся комментарии из оригинального поста
Похожие видео
0:30
Sensitive content
Qwen 3.8 Max beat Fable 5 at building 3D physics scenes for 7x cheaper! We gave two models the same task. Build three self-contained 3D scenes, each one HTML file with real physics that runs itself. Prompts: - A marble machine that lifts marbles up a wheel and drops them on a loop - A car factory assembly line - A sawmill cutting logs into planks Outputs: Qwen 3.8 Max: 46.6K tokens, $0.28 Fable 5: 38.7K tokens, $1.93 Qwen worked much harder on the details. Its textures and shadows look real. Fable 5 built a fast prototype that works and left it there. In the marble machine, Fable's marbles show up from nowhere at the top of the wheel. In Qwen's, you can see them go up. On the car line, Qwen's cars look much better and more real. Fable's sawmill is very strange. The saw looks wrong, and the log goes right through a beam before the saw cuts it. Atomic Chat will have day zero support to run Qwen 3.8 Max locally!
atomic.chat
210,569 просмотров • 21 дней назад

