LaGO: Latent Action Guidance for Online Reinforcement Learning
用潜在动作引导在线强化学习,让大模型从离线轨迹中提炼行动先验,训练更高效可控。
arXiv:2606.24669v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for planning and sequential decision-making, …