developer-tools
54 posts tagged developer-tools.
- When Documentation Doesn't Help Coding Agents: A Negative Result Worth Reading
- Drawgent puts a coding agent on an Excalidraw canvas: what a visual work surface actually changes
- Keeping programming enjoyable when LLMs write the first draft
- Qwen 27B on a 4090 is a useful reminder about local AI demos
- Claude Code’s verification loop is the real coding-agent primitive
- Ollaya points at a missing layer for decision models
- What One Developer Learned From a Month Without AI Tools
- Local tool-use evals are measuring your server too
- OpenAI’s Python SDK adds model handles before the launch story
- LangChain 1.4.2 fixes a small but real agent reliability problem
- Laya’s Jev claim needs a repo-first reality check
- What Two Boring openai-python Patches Reveal About Agent Reliability
- Jev is a decision model, not another chatbot
- LangChain’s typesafe alpha points at safer model routing
- The Skill Router You Already Have: Gavel Reads Routing From a Frozen LLM
- A vulnerability is not fixed because an AI bot saw it
- JPEG XL Is a Workflow Question, Not a Format War
- What llama.cpp v0.4.1 tells you about where local AI is heading
- A Filter That Strips AI Stories From Hacker News, and What It Reveals
- Devin testing its own work with GPT-6 Astra: what's real, what's reported
- Rune going open source is only the start of the diligence
- The RubyGems agent report is a supply-chain warning
- The useful part of a smaller LLM gateway is not the size
- OpenAI’s Python SDK gets key expiration controls, not just new image hooks
- OpenAI's Python SDK Gets More Honest About Agent Failures
- The real lesson in a 90% Claude Code token cut
- What LangChain Core 1.6.2 Fixes for Agent Builders
- Repo-distilled skills are the missing middle layer for research agents
- ACToR targets the tokens where repo-level code generation breaks
- When LLM memory becomes program analysis
- MCR-Bench shows code review agents still lose the plot
- Anthropic’s Python SDK is tightening the agent plumbing
- Prime Agent treats the harness as part of the model
- Codex over Claude is a workflow signal, not a verdict
- What LangChain's perplexity 1.4.1 patch says about agent plumbing
- What Actually Makes a Claude Code Session Productive
- What OpenAI's v3.1.0 SDK Changelog Tells Us About Its Roadmap
- RingCentral’s AI-native work story is really about the handoff
- VICBench shows vulnerability detection still needs humans
- AI security review is hitting Bitcoin repos, not just toy code
- AI coding cost control is an engineering workflow problem
- Kitesurf and the browser built for agents, not humans
- LLMs as semantic scouts for compiler optimizations
- Go’s generic collections proposal is really about shared defaults
- RL can teach code models to care about runtime, but the stopwatch is the hard part
- Half-Life 2 on HaikuOS and the AI runtime tax
- LangChain’s xAI update is really about model contracts
- Stack Overflow’s AI problem is the missing feedback loop
- Leanstral 1.5 points at proofs as a workflow, not a stunt
- A GPT-5.6 rumor matters only if Codex changes with it
- Agentic coding works best when the repo has rails
- The Model Got Smarter and My Tool Got Dumber
- Code-as-image is a cost hack, not a free lunch
- Anthropic’s Python SDK points to agents as infrastructure, not demos