fine-tuning
16 posts tagged fine-tuning.
- Teaching a Reasoning Model to Know When It's Sure Cuts Its Token Bill
- Translation Fine-Tuning Breaks the Controls General Evals Miss
- ComPO and the Case Against Optimizing the Loss You Wrote Down
- Distillation needs calibration when the teacher is biased
- What Fyxer's AI inbox assistant gets right about trust
- The Hallucination Detector That Doesn't Transfer to Your Domain
- The SFT-RL Split Has a Wide Safe Zone, and You Can Find It Cheap
- What LoRA Rank Actually Buys You, According to the Attention Math
- Reasoning fine-tunes can move the safety vector
- What the Summer 2026 Open Model Data Actually Shows
- ROPD treats poisoned fine-tunes as a distribution problem
- κ-LoRA makes LoRA tuning selective instead of uniform
- LoRA patches may outlive the base model update
- PPL-Factory makes the case for smaller fine-tuning sets
- CRAFT Turns Rubrics Into a Fine-Tuning Map
- When the Task Matches the Objective: What MTO Says About Fine-Tuning Small Models