ai-research
20 posts tagged ai-research.
- Alignment baked into pretraining, not bolted on later
- AutoDesign turns paper-to-poster into a harness the agent rewrites itself
- LLMs Know When to Back Off, But Still Guess Too Specifically
- OmniScientist argues that AI scientists need eyes, not just workflows
- CLAUDE.md bloat is a memory problem, not a prompt problem
- Concise answers can weaken reasoning in fused LLM training
- Two prepared policies may be the sweet spot for uncertain MDPs
- AI research agents can code, but they still can’t judge the work
- The Automation Ceiling Nobody Prices In: When Human Participation Is the Product
- No Best Harness: Automated Discovery Systems Fail to Generalize
- Uncertainty metrics should follow the loss, not the other way around
- Vision models are getting the scene right, but not always the gaze
- Leanstral 1.5 points at proofs as a workflow, not a stunt
- DiaLLM separates dialect understanding from dialect writing
- Language critiques are a better training signal than a score, if you can afford them
- LLMs brainstorm like synthesis machines, not researchers
- TraceLab shows coding agents are an infrastructure workload now
- Google's Paper Assistant Wants to Catch Your Math Errors Before a Reviewer Does
- Nash solvers have preferences when the value is identical
- AI math is less about genius than verification