Skip to content
{ ken ashe }
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
← Digest / Tags

// tag

distillation

15 posts tagged distillation.

  • UECR-GRPO treats the teacher as evidence, not an oracle Sep 24, 2026
  • Xiaomi's MiMo V2.6 Ships in Three Flavors, and the Split Matters Sep 22, 2026
  • When the Teacher Knows to Quit: RetireOPD and Distillation for Agents Sep 18, 2026
  • Distillation needs calibration when the teacher is biased Sep 16, 2026
  • OptiFlow treats offline RL policy learning as sample matching Sep 15, 2026
  • Garry Tan’s distillation argument is really about AI capability access Sep 14, 2026
  • On-policy distillation may need fewer prompts and better absorption Sep 4, 2026
  • On-policy distillation may be pruning tails, not teaching Sep 1, 2026
  • The token-budget bug hiding in multi-teacher distillation Aug 20, 2026
  • When the Teacher and the Verifier Disagree: Fixing On-Policy Distillation for Long Context Aug 20, 2026
  • A Strong Model Can Scaffold a Weak One Without Any Retraining Aug 13, 2026
  • Relay-OPD and the prefix failure problem in on-policy distillation Jul 29, 2026
  • OPD2 tries to distill reasoning by subtracting the base model Jul 17, 2026
  • Reusing RL Gains Across Model Sizes: The Case for Direct-OPD Jul 7, 2026
  • DemoPSD Treats Teacher Disagreement as a Training Signal Jul 3, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Explore

  • Building
  • Writing
  • Digest
  • Topics

About & Press

  • About
  • Newsroom
  • Media Kit
  • Lucky Domains

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public