Skip to content
arXiv cs.CL · Papers

MILES: Modular Instruction Memory with Learnable Selection for Self-Improving LLM Reasoning

arXiv:2607.06974v1 Announce Type: new Abstract: Large language models (LLMs) increasingly improve their reasoning at test time via additional computation, yet most existing works treat each problem in isolation. When problems arrive sequentially, accumulating reusable experience across them can further improve performa