we published a blog on hugging face at possibly the worst time yesterday lol
congrats to the HF team on the big news! 💚
here's a fine-tuning tutorial showing how to
• fine-tune a tiny LFM2.5-350M model
• in 100 GRPO steps using TRL
• for better structured outputs
blog:
colab: