MDRSS · MARKDOWN SNAPSHOT
9.8/10

LoRA & QLoRA Recipes

This skill assumes the routing decision already happened — finetuning-method-selection should have already pointed here because the data shape is demonstrations (SFT), not preference pairs or a verifiable reward signal. What follows is the current best-practice recipe for confi. Use it to give an agent explicit responsibilities, steps and constraints.

#llm-engineering / card #901370★ 0◌ 0snapshot 2026-08-04