local-ai
36 posts tagged local-ai.
- Free local AI models are not charity
- Prompt lookup drafting gets faster in llama.cpp, but the win is workload-specific
- Qwen 27B on a 4090 is a useful reminder about local AI demos
- Qwen architecture rumors do not make a 3090 fast by default
- Ling 3.0 Tiny makes old CPUs interesting again
- Swift Qwen’s speed claim is a local AI reminder, measure the whole loop
- Local tool-use evals are measuring your server too
- Xiaomi's MiMo V2.6 Ships in Three Flavors, and the Split Matters
- Mini-AGI makes local training the interesting part
- CXMT’s reported memory ramp matters more for inference than model hype
- Qwen-Image-2.1 puts open image editing closer to production work
- Treat the RX 10800 XT local AI claim as a software question, not a speed claim
- Mistral and Mozilla put AI in the browser: what 'private' actually means here
- What llama.cpp v0.4.1 tells you about where local AI is heading
- A thin local AI signal still tells builders what to watch
- Local LLMs are getting useful because constraints are back
- The Hugging Bay and the weak link in open model distribution
- Local LLMs as the first responder for a compromised PC
- WebGPU Kernels Move Local AI Closer to the Browser
- A 192GB Framework board would make local AI less cramped
- llama.cpp’s CPU backlog is a map of local AI’s next gains
- StemDeck Puts Stem Separation on Your Own Machine
- A 36-node DGX Spark homelab points at agent infrastructure, not just bigger inference
- r/LocalLLaMA and the quiet maturity of local AI
- Qwen thinking levels need task-level tests, not vibes
- A rumored 96GB RTX 5090 is a planning signal, not a purchase plan
- DeepSeek-V4-Flash on a 3090 shifts the bottleneck to DDR5
- DeepSeek-V4-Flash on a Mac is an I/O story, not a parameter-count story
- Kimi K3 on 8 GB RAM is a systems lesson, not a serving plan
- llama.cpp support is becoming the real local AI distribution layer
- Qwen’s next move matters after the weights land
- The Qwen3.8 rumor is really a VRAM planning signal
- Bonsai 27B makes local agents smaller, not magically smarter
- Bonsai 27B makes phone inference the claim to inspect
- Local AI is a computing right, not a hobbyist preference
- Local AI is insurance, not a bunker