
How LLM Reasoning Actually Works, and When It Is Not Worth Paying For
Extended thinking is not a smartness switch. It is extra sampled tokens, billed as output, trained by reinforcement learning on verifiable answers, and it has a measurable point past which accuracy goes down.



