fine-tuning-guide
Model fine-tuning covering dataset preparation, LoRA and QLoRA, instruction tuning, RLHF and DPO, benchmarking, overfitting prevention, compute requirements, Hugging Face Trainer, and the fine-tuning vs prompt engineering decision. Use when the user asks about fine tuning guide, fine tuning guide best practices, or needs guidance on fine tuning guide implementation. Do NOT use when the user needs a different specialized skill or is asking about an unrelated technology domain.