Lead story
Ask your AI
Top stories
Models & availability
Latest
Lead story
Ask your AI
Top stories
Models & availability
Latest
Hugging Face published an open-source recipe and notebook to fine-tune Liquid AI's LFM2.5-350M model using GRPO via the TRL library, improving structured-output compliance from 22.6% to 29.7% on the IFStruct benchmark. The guide runs on a free-tier Colab or Kaggle GPU using 500 samples and 100 training steps. Materials are available on GitHub.
From the source
The full run takes around 500 samples and 100 training steps, small enough for a free-tier Colab or Kaggle GPU, and is available on GitHub.
huggingface.co