reasoning-models
11 posts tagged reasoning-models.
- BDH-CQ Makes ARC Reasoning Cheaper by Thinking in Latent Space
- Concise answers can weaken reasoning in fused LLM training
- What 'Test-Time Scaling' Actually Means When You Read a Benchmark
- AI reasoning can look right while taking shortcuts
- OpenAI's Ten Math Results: What Counts as a Real Advance
- PPL-Factory makes the case for smaller fine-tuning sets
- Chess Shows What RL Actually Does to a Reasoning Model
- OPD2 tries to distill reasoning by subtracting the base model
- AdaPrefix-GRPO turns hard reasoning failures into training signal
- Agon Grades the Reasoning, Not Just the Answer
- Reasoning Traces as a Difficulty Sensor: What Epi2Diff Gets Right