Skip to content
arXiv cs.CL · Papers

DORA Explorer: Improving the Exploration Ability of LLMs Without Training

arXiv:2604.17244v2 Announce Type: replace Abstract: Large language model (LLM) agents for sequential decision-making struggle to produce diverse outputs. This leads to insufficient exploration, suboptimal solutions, and repeated actions. Actions are generated at the sequence level, but existing sampling strategies, suc