we published a blog on hugging face at possibly the worst time yesterday lol
congrats to the HF team on the big news! ๐
here's a fine-tuning tutorial showing how to
โข fine-tune a tiny LFM2.5-350M model
โข in 100 GRPO steps using TRL
โข for better structured outputs
blog:
colab: