Loading video...

Video Failed to Load

Go Home

Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub:

2,252,490 views • 6 days ago •via X (Twitter)

56 Comments

White Circle's profile picture
White Circle6 days ago

Every model we train at White Circle now uses Halo. The same codebase runs LoRA on a 24 GB GPU, multi-node training on B300s, and async RL. Technical blog:

White Circle's profile picture
White Circle6 days ago

Other training frameworks often require a separate implementation for each model family. In Halo, adding a new model family takes just 100 lines of connecting a wrapper instead of rewriting the model.

White Circle's profile picture
White Circle6 days ago

One YAML file contains the model, training method, GPU setup and checkpoint settings. One command starts training.

White Circle's profile picture
White Circle6 days ago

Halo supports data, tensor, context, expert and expert-tensor parallelism. It also includes optimized attention, grouped-GEMM and fused loss kernels.

White Circle's profile picture
White Circle6 days ago

Halo is not limited to supervised fine-tuning. It supports async RL and training with external environments. We are excited to partner with the @sgl_project team to make it the primary engine for rollouts.

White Circle's profile picture
White Circle6 days ago

On @OpenAI gpt-oss-20b, Halo delivered 2.3–2.8x the throughput of stock TRL with less peak memory. Both runs used the same FlashAttention, Liger, fused cross-entropy and grouped-GEMM optimizations.

White Circle's profile picture
White Circle6 days ago

We partnered with @liquidai to add native Halo support for LFM2.5-8B-A1B. On one B300, Halo was up to 20% faster than next best open-source framework.

White Circle's profile picture
White Circle6 days ago

We also used Halo to fine-tune @Zai_org GLM-4.7-Flash on 177M tokens of agentic traces. The resulting model improved SWE-rebench-V2 by @nebiusai from 33% to 42%. Halo reached up to 1.63× TRL throughput on the same model precision and data. Model and write-up:

White Circle's profile picture
White Circle6 days ago

Thanks for reading to the end! Star us on GitHub: Read more:

Daniel Smidstrup's profile picture
Daniel Smidstrup6 days ago

Open source is the way, and it's only going to get more and more relevant in the coming time! :)

White Circle's profile picture
White Circle6 days ago

True

leo's profile picture
leo6 days ago

insane way to kick off the week -- congrats @whitecircle team

White Circle's profile picture
White Circle6 days ago

Lets go!

neural nets.'s profile picture
neural nets.6 days ago

crazy reading through the docs

Vishal Singh 🥑's profile picture
Vishal Singh 🥑6 days ago

love that this is open source, congrats team ❤️

Catalin's profile picture
Catalin6 days ago

Congrats on the awesome work, team! Good to see training infra being open source 👏 Starred the repo to contribute to the growth.

White Circle's profile picture
White Circle6 days ago

thanks!

Paul Mit's profile picture
Paul Mit6 days ago

great idea you've got my star, guys

giulia's profile picture
giulia6 days ago

took me one read to understand exactly who this is for. that almost never happens with infra launches. congrats team!!! starred 👀

Mac's profile picture
Mac6 days ago

Peak improvement, I will try it from first hand

sui's profile picture
sui6 days ago

so we get Grok 4.7 along with this today? nice.

White Circle's profile picture
White Circle6 days ago

Crazy day!

Maxime Labonne's profile picture
Maxime Labonne6 days ago

Congrats guys!

Virgile RIETSCH's profile picture
Virgile RIETSCH6 days ago

This is huge !!

Levan Kvirkvelia's profile picture
Levan Kvirkvelia6 days ago

haloshi framework

𝗕𝗿𝗶𝗮𝗻 𝗥𝗲𝘆's profile picture
𝗕𝗿𝗶𝗮𝗻 𝗥𝗲𝘆6 days ago

wait what this is incredible

White Circle's profile picture
White Circle6 days ago

ty Brian! show the repo some love ⭐

ivy's profile picture
ivy6 days ago

so excited!!!

White Circle's profile picture
White Circle6 days ago

🥰

Ted | unfair.so's profile picture
Ted | unfair.so6 days ago

Congrats team!

Soraia's profile picture
Soraia6 days ago

Great work! You got my star ⭐️

Nidhi Singh's profile picture
Nidhi Singh6 days ago

love that this is open source, congrats!

Rushil Chopra's profile picture
Rushil Chopra6 days ago

I have always wanted to train a custom math model, let’s see how this goes!

White Circle's profile picture
White Circle6 days ago

Keep us posted, any feedback is welcome

AshutoshShrivastava's profile picture
AshutoshShrivastava6 days ago

love seeing more training infrastructure get open-sourced. congrats on the release!

Ion's profile picture
Ion6 days ago

Congrats! Impressive what you delivered there 👀

Nikita Andersson's profile picture
Nikita Andersson6 days ago

@sarapenrique This is fabulous

Hauke 🌱🦀's profile picture
Hauke 🌱🦀6 days ago

Uh training my own models for my products would be pretty neat. I need better hardware 👀

White Circle's profile picture
White Circle6 days ago

💚

Prasenjit's profile picture
Prasenjit6 days ago

oh this is sick actually, well cooked

Kalash's profile picture
Kalash6 days ago

banger stuff 💥

Dan Kulkov's profile picture
Dan Kulkov6 days ago

insane start of the week congrats team!

Aditi's profile picture
Aditi6 days ago

wait… 2.8x faster than stock trl ??

Vasko's profile picture
Vasko6 days ago

congrats to the team!

Lucas Valbuena's profile picture
Lucas Valbuena6 days ago

congrats!

White Circle's profile picture
White Circle6 days ago

Thank you

Sameer's profile picture
Sameer6 days ago

congrats! def keeping an eye on this one.

Suraj Sharma's profile picture
Suraj Sharma6 days ago

The interesting part is keeping models in native HuggingFace format while getting that throughput boost.

White Circle's profile picture
White Circle6 days ago

Worked hard on this

Suraj Sharma's profile picture
Suraj Sharma6 days ago

LFG 🫡🙌

Jackson's profile picture
Jackson6 days ago

Really cool. Really unique and impressive numbers :•)

Marco Franzon's profile picture
Marco Franzon6 days ago

The only limit now is the hardware. Everyone can now train or finetune a model without rewriting it for the training phase. Awesome!

White Circle's profile picture
White Circle6 days ago

thanks Marco!

Marco Franzon's profile picture
Marco Franzon6 days ago

Great work !

Deni's profile picture
Deni6 days ago

Congrats on the launch team! Starred.

Ole Lehmann's profile picture
Ole Lehmann6 days ago

happy launch day guys"

Related Videos