arXiv 2604.10739v1 investigates the "over-thinking" phenomenon in test-time compute: when given a simple problem, reasoning models still generate thousands of tokens of "thinking" before answering, with the extra compute not improving and sometimes hurting accuracy. The paper proposes methods to detect and skip unnecessary thinking steps.