HF Quantization

community
Activity Feed

AI & ML interests

Making AI models smaller, and faster with Quantization

medmekk 
posted an update 1 day ago
view post
Post
2030
🚀 Introducing Halo 1.0

Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle.

It comes with:
🧠 Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillation
🤖 Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool use
⚡ ~2.8× TRL throughput on 8× B300 (EP+FSDPv2, FA4, fp8/fp4)
🤗 Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4…)
🛠️ One halo command, prebuilt Docker images, and docs for humans and agents

💻 https://github.com/whitecircle/halo

Try it and tell us what you're training
  • 1 reply
·