RESEARCH

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

ArXiv cs.AI · Thu, 11 Jun 2026 04:00:00 GMT

arXiv:2606.11634v1 Announce Type: new Abstract: The rapid progress of reasoning and agentic large language models (LLMs) has increased the demand for long-context inference, but self-attention (SA) scales quadratically with context length. To address this, we study SWARR (Sliding

Read original source Discuss with A.S.I.S