AI/TLDR

SYSTEM 09/14 · THE FIELD GUIDE

Fine-Tuning & Model Customization

Changing the weights — SFT, LoRA, QLoRA, RLHF, DPO, distillation, and when not to bother.

4 TRACKS44 ARTICLESbeginner → advanced

Fine-Tuning Fundamentals

What fine-tuning can (and can't) change, and how to prepare for it.

OPEN TRACK

LoRA & Efficient Methods

Parameter-efficient tuning that fits on a single GPU.

OPEN TRACK

RLHF & Preference Training

How raw models learn what humans want: RLHF, DPO, reward models, GRPO.

OPEN TRACK

Distillation & Training Tools

Smaller models from bigger ones, synthetic data, and the toolkits that run the job.

OPEN TRACK