# #trl — MDRSS hashtag feed

> Public MDRSS cards tagged #trl.
> Canonical feed: https://mdrss.com/feeds/trl

## Cards (2)

### [Unsloth ↔ TRL/PEFT Mapping](https://mdrss.com/llm-engineering/training-and-fine-tuning/901372/901372.md)

Unsloth is a fast-kernel wrapper over PEFT and TRL, not a replacement API — every Unsloth kwarg below has a plain TRL/PEFT equivalent. Use this table to translate an Unsloth config to plain TRL (or back), and to know which knob lives on which object in the current TRL API. Use it to give an agent explicit responsibilities, steps and constraints.

Classification: llm-engineering/training-and-fine-tuning · Feed: llm-engineering · Updated: 2026-08-04T13:54:51.641Z · Version: 1

### [TRL - Transformers Reinforcement Learning](https://mdrss.com/llm-engineering/inference-and-quantization/901071/901071.md)

TRL is a cutting-edge library designed for post-training foundation models using advanced techniques like Supervised Fine-Tuning (SFT), Group Relative Policy Optimization (GRPO), and Direct Preference Optimization (DPO). Built on top of the 🤗 Transformers ecosystem, TRL supports. Use it to navigate the topic and choose relevant methods, papers or tools.

Classification: llm-engineering/inference-and-quantization · Feed: llm-engineering · Updated: 2026-08-04T13:54:51.641Z · Version: 1
