Skip to content
{ ken ashe }
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
  • Building
  • Writing
  • About
  • Newsroom
  • Digest
← Digest / Tags

// tag

local-ai

36 posts tagged local-ai.

  • Free local AI models are not charity Sep 27, 2026
  • Prompt lookup drafting gets faster in llama.cpp, but the win is workload-specific Sep 27, 2026
  • Qwen 27B on a 4090 is a useful reminder about local AI demos Sep 27, 2026
  • Qwen architecture rumors do not make a 3090 fast by default Sep 27, 2026
  • Ling 3.0 Tiny makes old CPUs interesting again Sep 26, 2026
  • Swift Qwen’s speed claim is a local AI reminder, measure the whole loop Sep 26, 2026
  • Local tool-use evals are measuring your server too Sep 23, 2026
  • Xiaomi's MiMo V2.6 Ships in Three Flavors, and the Split Matters Sep 22, 2026
  • Mini-AGI makes local training the interesting part Sep 21, 2026
  • CXMT’s reported memory ramp matters more for inference than model hype Sep 20, 2026
  • Qwen-Image-2.1 puts open image editing closer to production work Sep 20, 2026
  • Treat the RX 10800 XT local AI claim as a software question, not a speed claim Sep 20, 2026
  • Mistral and Mozilla put AI in the browser: what 'private' actually means here Sep 16, 2026
  • What llama.cpp v0.4.1 tells you about where local AI is heading Sep 14, 2026
  • A thin local AI signal still tells builders what to watch Sep 13, 2026
  • Local LLMs are getting useful because constraints are back Sep 13, 2026
  • The Hugging Bay and the weak link in open model distribution Sep 13, 2026
  • Local LLMs as the first responder for a compromised PC Sep 6, 2026
  • WebGPU Kernels Move Local AI Closer to the Browser Sep 1, 2026
  • A 192GB Framework board would make local AI less cramped Aug 30, 2026
  • llama.cpp’s CPU backlog is a map of local AI’s next gains Aug 30, 2026
  • StemDeck Puts Stem Separation on Your Own Machine Aug 29, 2026
  • A 36-node DGX Spark homelab points at agent infrastructure, not just bigger inference Aug 23, 2026
  • r/LocalLLaMA and the quiet maturity of local AI Aug 23, 2026
  • Qwen thinking levels need task-level tests, not vibes Aug 22, 2026
  • A rumored 96GB RTX 5090 is a planning signal, not a purchase plan Aug 9, 2026
  • DeepSeek-V4-Flash on a 3090 shifts the bottleneck to DDR5 Aug 2, 2026
  • DeepSeek-V4-Flash on a Mac is an I/O story, not a parameter-count story Aug 2, 2026
  • Kimi K3 on 8 GB RAM is a systems lesson, not a serving plan Aug 2, 2026
  • llama.cpp support is becoming the real local AI distribution layer Aug 2, 2026
  • Qwen’s next move matters after the weights land Jul 19, 2026
  • The Qwen3.8 rumor is really a VRAM planning signal Jul 19, 2026
  • Bonsai 27B makes local agents smaller, not magically smarter Jul 15, 2026
  • Bonsai 27B makes phone inference the claim to inspect Jul 15, 2026
  • Local AI is a computing right, not a hobbyist preference Jul 4, 2026
  • Local AI is insurance, not a bunker Jun 28, 2026

Ken Ashe ·AI application builder ·CPA ·PMP

Building with AI in public. No hype, no doom. Receipts only.

hello@kenashe.ai

Explore

  • Building
  • Writing
  • Digest
  • Topics

About & Press

  • About
  • Newsroom
  • Media Kit
  • Lucky Domains

Social

  • LinkedIn
  • X
  • GitHub
  • RSS

Legal

  • Privacy
  • Terms
  • Disclosure

© 2026 Ken Ashe ·Built with AI in public