Introducing halo - HF format trainer with up to 3x the speed of TRL and all parallelisms
i work on Halo at White Circle. we built it after repeatedly hitting the gap between stock HF/TRL setups and frameworks that require a separate model implementation and checkpoint format. Halo adds expert, context, tensor and expert-tensor parallelism directly to Hugging Face models. trainers remai…
Read the full story at r/huggingface ↗
Timeline · 1 report
- 2026-09-21 17:41 · r/huggingface
Introducing halo - HF format trainer with up to 3x the speed of TRL and all parallelisms