Loading video...
Video Failed to Load
Red Hat AI just shipped DFlash speculator checkpoints for two of NVIDIA AI's most powerful open models: → Nemotron Ultra 550B → Nemotron Super 120B On math and reasoning: ~5 out of 7 draft tokens accepted on average. On code (HumanEval): ~3.4 out of 7. Both checkpoints trained with... show more
15,077 views • 2 months ago •via X (Twitter)
4 Comments

Helen Zhao2 months ago
@NVIDIAAI Training compute for this model was generously provided by @LambdaAPI, a leading cloud platform built for AI training and inference. Thanks again @TheZachMueller for all the help and support along the way!

RiftStack AI2 months ago
@NVIDIAAI any intuition why code accepts so many fewer draft tokens than math?

Oli Wilkins2 months ago
@NVIDIAAI How does this compare to something like EAGLE?

Mike Gannotti2 months ago
@NVIDIAAI I’m really looking forward to seeing what NVIDIA drops when Nemotron hits 4. 3 is already such a strong set of models


