註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

shirish
@shiri_shh
Generalist | AI | Tech | Startups ▪︎ Locked in
加入 December 2014
1.5K 正在關注    45.7K 粉絲
btw fine-tuning open-source models at 2x the speed is easy now. you just need: - your favorite base model (Llama, Qwen, Mistral) - a single YAML config file - Halo from @whitecircle up to 2.8x faster throughput, lower peak VRAM, and 0 checkpoint conversion scripts
顯示更多
Introducing Halo, the best framework for post-training of open-source models. Halo delivers up to 2.8x the throughput of stock TRL with less peak memory, while models stay in their native HuggingFace format. Star us on GitHub:
顯示更多