SAGE: Mitigating Long-Horizon Reasoning Biases via Topological Guidance
arXiv:2609.30192v1 Announce Type: new Abstract: Long-horizon reasoning remains a central challenge for large language models (LLMs) under sparse-reward regimes. We argue that this brittleness arises from two biases induced by complex reasoning spaces: an exploration bias, where models are drawn…