注册并分享邀请链接,可获得视频播放与邀请奖励。

shirish
@shiri_shh
Generalist | AI | Tech | Startups ▪︎ Locked in
加入 December 2014
1.5K 正在关注    45.7K 粉丝
btw fine-tuning open-source models at 2x the speed is easy now. you just need: - your favorite base model (Llama, Qwen, Mistral) - a single YAML config file - Halo from @whitecircle up to 2.8x faster throughput, lower peak VRAM, and 0 checkpoint conversion scripts
显示更多
Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub:
显示更多