Fine-tunes language models with TRL using methods like SFT and DPO.
/trl-traininghuggingface/skills
/trl-training