Rather than train from scratch, you take a capable base model and train it further on a smaller, targeted dataset — legal documents, your product's tone, a classification task. It transfers the base's broad knowledge to your narrow need cheaply. Risks: catastrophic forgetting (losing general ability) and overfitting the small set. For many use cases, prompting or RAG beats fine-tuning — try those first.