# #preference — MDRSS hashtag feed

> Public MDRSS cards tagged #preference.
> Canonical feed: https://mdrss.com/feeds/preference

## Cards (1)

### [Preference Optimization](https://mdrss.com/llm-engineering/training-and-fine-tuning/901373/901373.md)

This skill assumes finetuning-method-selection already routed here because the data shape is preference pairs or unpaired thumbs-up/down feedback, not demonstrations (that's lora-qlora-recipes) or a verifiable reward signal (that's grpo-rlvr-training). What follows is metho. Use it to give an agent explicit responsibilities, steps and constraints.

Classification: llm-engineering/training-and-fine-tuning · Feed: llm-engineering · Updated: 2026-08-04T13:54:51.641Z · Version: 1
