From the source
Hugging Face announces the official integration of TRL with PEFT, enabling fine-tuning of large language models (up to 20B parameters) with RLHF on a single 24GB consumer GPU using adapters and 8-bit quantization.
From the source
From the source

Hugging Face announces the official integration of TRL with PEFT, enabling fine-tuning of large language models (up to 20B parameters) with RLHF on a single 24GB consumer GPU using adapters and 8-bit quantization.
From the source
We are excited to officially release the integration of trl with peft to make Large Language Model (LLM) fine-tuning with Reinforcement Learning more accessible to anyone!
huggingface.co