llama-cpp
9 posts tagged llama-cpp.
- Prompt lookup drafting gets faster in llama.cpp, but the win is workload-specific
- Ling 3.0 Tiny makes old CPUs interesting again
- Treat the RX 10800 XT local AI claim as a software question, not a speed claim
- What llama.cpp v0.4.1 tells you about where local AI is heading
- llama.cpp’s CPU backlog is a map of local AI’s next gains
- Local agentic coding at 60 tokens per second is only half the test
- DeepSeek-V4-Flash on a 3090 shifts the bottleneck to DDR5
- llama.cpp support is becoming the real local AI distribution layer
- What llama.cpp's commit log tells us about AI-written code