Skip to content
{ ken ashe }
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
← Digest / Tags

// tag

model-evaluation

12 posts tagged model-evaluation.

  • Pachocki’s warning points to safety gates, not slower vibes Sep 8, 2026
  • Uncensored Qwen edits show why model cards are not enough Sep 6, 2026
  • Certified world models still have blind topology Aug 31, 2026
  • CAST makes clinical model audits more concrete Aug 28, 2026
  • Qwen thinking levels need task-level tests, not vibes Aug 22, 2026
  • TokEval makes tokenizer choice measurable before pretraining Aug 19, 2026
  • EPC scores explanations by testing what the model can lose Aug 3, 2026
  • Instruction following is the local model test benchmarks miss Aug 2, 2026
  • Vacuum 16T turns model size into a metadata bug Aug 2, 2026
  • Tabular foundation models still stumble when the rows change Jul 29, 2026
  • Uncertainty metrics should follow the loss, not the other way around Jul 17, 2026
  • New model releases do not reset the advantage Jul 16, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Explore

  • Building
  • Writing
  • Digest
  • Topics

About & Press

  • About
  • Newsroom
  • Media Kit
  • Lucky Domains

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public