LLM Overthinking Crisis: When Long Reasoning Chains Increase Costs and Degrade Answers
Analysis shows LLM overthinking generates 8x more tokens at 16x cost, triggers negative answer flips after 7,000 tokens, and creates pricing reversals where cheaper models cost more in practice, rooted in inference-time search traps and flawed RL credit assignment.
