1
DORA Explorer: Improving the Exploration Ability of LLMs Without Training
不训练也能让大模型更会“探索”?DORA Explorer 提出全新推理时方法,显著增强LLM在复杂任务中的决策能力与探索广度。
arXiv:2604.17244v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents for sequential decision-making struggle to produce diverse…