mixeden
ยท
AI & ML interests
None yet
Recent Activity
reacted to medmekk's post with ๐ about 2 hours ago ๐ Introducing Halo 1.0
Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle.
It comes with:
๐ง Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillation
๐ค Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool use
โก ~2.8ร TRL throughput on 8ร B300 (EP+FSDPv2, FA4, fp8/fp4)
๐ค Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4โฆ)
๐ ๏ธ One halo command, prebuilt Docker images, and docs for humans and agents
๐ป https://github.com/whitecircle/halo
Try it and tell us what you're training View all activity Organizations
published an article over 1 year ago view article CircleGuardBench: New Standard for Evaluating AI Moderation Models
whitecircle
โข โข 59