Loading video...
Video Failed to Load
We open-sourced QeRL — Quantization-enhanced Reinforcement Learning ! 🧠 4-bit quantized RL training 💪 Train a 32B LLM on a single H100 GPU ⚙️ 1.7× faster overall training 🎯 Accuracy on par with bfloat16-level accuracy 🔥 Supports NVFP4 quantization format Moreover, we show that quantization helps exploration in RL... show more
69,720 views • 7 months ago •via X (Twitter)
0 Comments
No comments available
Comments from the original post will appear here
