From 4bc79bbfb269491b0fee051f15a78f33db47868b Mon Sep 17 00:00:00 2001 From: Vinta Chen Date: Thu, 1 Oct 2026 20:45:10 +0800 Subject: [PATCH] feat: add trl to AI and Agents > Fine-tuning Co-Authored-By: Claude --- README.md | 1 + 1 file changed, 1 insertion(+) diff --git a/README.md b/README.md index 4035871e..8cf20e0c 100644 --- a/README.md +++ b/README.md @@ -173,6 +173,7 @@ _Libraries for building AI applications, LLM integrations, and autonomous agents - [diffusers](https://github.com/huggingface/diffusers) - A library that provides pre-trained diffusion models for generating and editing images, audio, and video. - Fine-tuning - [peft](https://github.com/huggingface/peft) - A library for parameter-efficient fine-tuning of large pretrained models. + - [trl](https://github.com/huggingface/trl) - A library for post-training transformer language models with SFT, DPO, GRPO, and other trainers. - [unsloth](https://github.com/unslothai/unsloth) - Faster, lower-memory LLM fine-tuning, as a Python library or a desktop app. - [axolotl](https://github.com/axolotl-ai-cloud/axolotl) - A framework for fine-tuning and post-training large language models. - Speech