ThinkReset: Learnable Intermediate Interface Construction for Bounded-Context Long-Horizon Reasoning

arXiv:2607.28642v1 Announce Type: new Abstract: Long chain-of-thought reasoning improves performance on complex problems, but it also introduces redundancy accumulation, context overflow, and error anchoring. We argue that under bounded context windows, the core bottleneck is not trajectory compression or test-time control, but the absence of a reusable intermediate interface that can replace discarded history and support continued solving. We further identify a key failure mode of outcome-rewar...

arXiv cs.AI ·Fei Ding, Yongkang Zhang, Runhao Liu, Yuhao Liao, Zijian Zeng ·
compartilhar: