Skip to content
{ ken ashe }
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
← Digest / Tags

// tag

alignment

15 posts tagged alignment.

  • What LLMs Miss About Haitian Creole, and Why Low-Resource Culture Breaks Evals Sep 28, 2026
  • onPanda turns alignment feedback into token-level steering Sep 22, 2026
  • DiaVLo Turns Vision-Language Model Failures Into Named Behaviours Sep 21, 2026
  • Toxicity Scores Can Miss Sanitized Bias in GPT Outputs Sep 18, 2026
  • ComPO and the Case Against Optimizing the Loss You Wrote Down Sep 17, 2026
  • The Only Safe Pace for AI Is Everyone Else's Sep 13, 2026
  • When Agents Lie to Pass the Eval Sep 13, 2026
  • SPINE Shows Sycophancy Gets Worse When Users Keep Pushing Sep 9, 2026
  • OpenAI’s alignment note is really about operational discipline Sep 7, 2026
  • What a Moral Probe Finds Inside an LLM: Structure, Not a Single Switch Aug 28, 2026
  • Alignment baked into pretraining, not bolted on later Aug 14, 2026
  • IMPFM turns online alignment into a particle swarm Jul 2, 2026
  • RLVR needs a taste model, not just a grader Jul 2, 2026
  • When Playing It Safe Makes Reward Hacking Worse Jun 30, 2026
  • Agent safety belongs outside the agent process Jun 25, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Explore

  • Building
  • Writing
  • Digest
  • Topics

About & Press

  • About
  • Newsroom
  • Media Kit
  • Lucky Domains

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public