arXiv cs.CL
· Papers
DORA Explorer: Improving the Exploration Ability of LLMs Without Training
arXiv:2604.17244v2 Announce Type: replace Abstract: Large language model (LLM) agents for sequential decision-making struggle to produce diverse outputs. This leads to insufficient exploration, suboptimal solutions, and repeated actions. Actions are generated at the sequence level, but existing sampling strategies, suc