Skip to content
{ ken ashe }
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
← Digest / Tags

// tag

coding-agents

36 posts tagged coding-agents.

  • When Documentation Doesn't Help Coding Agents: A Negative Result Worth Reading Sep 28, 2026
  • Drawgent puts a coding agent on an Excalidraw canvas: what a visual work surface actually changes Sep 27, 2026
  • Keeping programming enjoyable when LLMs write the first draft Sep 27, 2026
  • Claude Code’s verification loop is the real coding-agent primitive Sep 26, 2026
  • What One Developer Learned From a Month Without AI Tools Sep 26, 2026
  • SWE-Flux: The Benchmark That Asks If Coding Models Can Predict What Code Actually Does Sep 24, 2026
  • CliffCompaction makes long coding runs cheaper by refusing to summarize Sep 23, 2026
  • Coding agents overclaim when their work is incomplete Sep 18, 2026
  • The Harness Matters as Much as the Model in Coding Agents Sep 18, 2026
  • When a Coding Agent Drives a Robot, Task Success Isn't Safety Sep 18, 2026
  • Real-SWE and the Case for Testing Coding Agents on Code They Have Never Seen Sep 13, 2026
  • Devin testing its own work with GPT-6 Astra: what's real, what's reported Sep 12, 2026
  • Perplexity Handing GPT-6 Astra Production Access: What OpenAI's Claim Actually Means Sep 12, 2026
  • Codex as a lab scout for antimicrobial search Sep 10, 2026
  • OpenAI Says Its Own Researchers Now Lean on Coding Agents. What Does the Data Actually Show? Sep 6, 2026
  • Rust vtables are where AI code reviews get vague Sep 6, 2026
  • The real lesson in a 90% Claude Code token cut Sep 5, 2026
  • Terminal-Universe turns code-agent logs into reusable sandboxes Sep 4, 2026
  • A Minecraft clone is a weak coding-agent test Aug 30, 2026
  • GLM, Qwen, and the messy reality of visual coding agents Aug 30, 2026
  • Terminal-Bench 4.0 and the eval gap for smaller coding agents Aug 29, 2026
  • SWE-Prime argues coding agents need cleaner wins, not more wins Aug 28, 2026
  • Coding agents still struggle with whole-repo migrations Aug 25, 2026
  • Codex over Claude is a workflow signal, not a verdict Aug 23, 2026
  • Local agentic coding at 60 tokens per second is only half the test Aug 22, 2026
  • Qwen3.8-27B looks useful as overnight local coding labor Aug 16, 2026
  • Qwen’s BASIC ray-tracer demo is really about closed-loop coding Aug 16, 2026
  • Oracle’s OpenJDK AI-code ban is really about provenance Aug 8, 2026
  • Where multimodal embeddings and collaborative coding agents actually stand Aug 4, 2026
  • Instruction following is the local model test benchmarks miss Aug 2, 2026
  • MindForge trains coding agents on blank-repo software work Jul 30, 2026
  • Coding agents are becoming lab infrastructure Jul 29, 2026
  • Claude Fable 5 browser-game demos are really coherence tests Jul 5, 2026
  • TestEvo-Bench moves coding-agent evals closer to real maintenance Jul 3, 2026
  • TraceLab shows coding agents are an infrastructure workload now Jun 30, 2026
  • Coding agents need routers before they need bigger models Jun 27, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Explore

  • Building
  • Writing
  • Digest
  • Topics

About & Press

  • About
  • Newsroom
  • Media Kit
  • Lucky Domains

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public