Tag
1 post
SFT, LoRA, DPO, and RLHF all change a model's behavior — but none of them give it current knowledge. Here's how to decide between fine-tuning and RAG, why they're usually both needed, and what a complete production architecture looks like end to end.