← The reading
The bibliography
§5 · Reasoning and Test-Time Compute
Reasoning book ch4–5
CAPTURED
15 On the list1 Starred0 In the Atlas15 To read
Rungs
Where this section moves a number- Build a Reasoning Model From ScratchOff the ladder
The pick
The source author's must-read for this sectionThe papers
15 papers- PaCoRe: Learning to Scale Test-Time Compute with Parallel Coordinated Reasoning ↗
- Learning to Reason in 13 Parameters ↗
- InftyThink+: Effective and Efficient Infinite-Horizon Reasoning via Reinforcement Learning ↗
- Learning to Self-Verify Makes Language Models Better Reasoners ↗
- Does Your Reasoning Model Implicitly Know When to Stop Thinking? ↗
- Reasoning Models Struggle to Control Their Chains of Thought ↗
- Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs ↗
- Nemotron-Cascade 2: Post-Training LLMs with Cascade RL and Multi-Domain On-Policy Distillation ↗
- Test-Time Scaling Makes Overtraining Compute-Optimal ↗★
- Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets ↗
- When Can LLMs Learn to Reason with Weak Supervision? ↗
- AI Co-Mathematician: Accelerating Mathematicians with Agentic AI ↗
- LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling ↗
- Achieving Gold-Medal-Level Olympiad Reasoning via Simple and Unified Scaling ↗
- Share More, Search Less: Collaborative Parallel Thinking for Efficient Test-Time Scaling ↗
What counts as read
A paper is on this list because someone worth reading put it there. That is a pointer, not a claim: it counts as read only once it has a close-read file in kb/<topic>/raw/, which is what a close-read link on a row means. There is deliberately nothing to tick off here — the study desk is the only writer of study state, and a second way to mark something done is a second version of the truth.