Training Proactive and Personalized LLM Agents
让大模型从“被动应答”走向“主动服务”,教你训练会察言观色的个性化AI代理,附完整方法框架。
arXiv:2511.02208v2 Announce Type: replace Abstract: Despite rapid progress, current AI agents are primarily optimized for isolated task completion. We…
让大模型从“被动应答”走向“主动服务”,教你训练会察言观色的个性化AI代理,附完整方法框架。
arXiv:2511.02208v2 Announce Type: replace Abstract: Despite rapid progress, current AI agents are primarily optimized for isolated task completion. We…
从理论层面拆解事后辩论评判机制,给AI对齐与评估研究提供全新视角。
arXiv:2608.19002v1 Announce Type: new Abstract: Debates have recently emerged as a useful methodology for agentic AI to improve performance as well as…
多模态大模型如何量化不确定性?这项研究为高风险场景下的可靠决策提供新思路。
arXiv:2608.17084v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly answer questions whose correctness depends on vi…
首个面向AI-RAN的主动式KV缓存迁移框架,直击LLM推理中长序列与多边缘节点的时延痛点,架构设计值得关注。
arXiv:2608.16477v1 Announce Type: new Abstract: AI-RAN brings large language model (LLM) serving close to mobile users, but cellular handover can sepa…
让LLM智能体在动态威胁中自我进化防御,告别手工安全规则,安全研究新范式。
arXiv:2608.12977v1 Announce Type: cross Abstract: The expanding operational capabilities of large language model (LLM) agents introduce sophisticated …
LLM策略短板明显,记忆增强代理却能让推理能力飙升,值得关注。
arXiv:2608.12626v1 Announce Type: cross Abstract: Strategic reasoning in Large Language Models (LLMs) within long-horizon environments is often limite…
用世界模型为自动研究代理注入规划能力,探索智能体规模化的前沿路径。
arXiv:2608.12564v1 Announce Type: new Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResea…
用prefill激活值重思LLM路由选择,为多模型推理调度开辟高效新路径。
arXiv:2603.20895v3 Announce Type: replace-cross Abstract: Existing routers rely on semantic query features or handcrafted features, which often fail t…
不训练也能让大模型更会“探索”?DORA Explorer 提出全新推理时方法,显著增强LLM在复杂任务中的决策能力与探索广度。
arXiv:2604.17244v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents for sequential decision-making struggle to produce diverse…
让农业智能不再是巨头专属,一份开源农业AI系统的研究论文,为精准农业提供可落地的技术路径。
arXiv:2506.04571v3 Announce Type: replace Abstract: Agriculture is undergoing a major transformation driven by artificial intelligence (AI), machine l…
提出广义线性马尔可夫决策过程统一框架,为强化学习理论分析提供新视角,值得关注。
arXiv:2506.00818v2 Announce Type: replace-cross Abstract: Offline reinforcement learning for longitudinal studies often faces two linked challenges: r…
给大模型智能体装上“时间线图记忆”,用证据支撑时序推理,值得AI研究者一读
arXiv:2608.08055v1 Announce Type: new Abstract: Large language model (LLM) agents that assist users over weeks of conversation must remember what is c…
让编程智能体实时纠偏,在线监控与纠正引导新框架,大幅提升代码生成可靠性。
arXiv:2608.06701v1 Announce Type: cross Abstract: Fixing GitHub issues in large-scale projects is a long-horizon task, especially when a fix requires …
百万用户与AI谈恋爱,这项研究揭开AI伴侣应用的隐藏风险
arXiv:2605.08093v2 Announce Type: replace-cross Abstract: The use of chatbots for various forms of companionship is growing rapidly, raising a myriad …
一套端到端的AI智能体审计引擎,为Agent行为安全与合规治理提供系统化审查方案
arXiv:2608.07346v1 Announce Type: new Abstract: With the rapid advancement of large language models (LLMs), harnesses have become essential infrastruc…
从工具调用看AI发展的长期主义:通用能力终将胜过精巧设计,值得每位关注智能体的读者深思。
arXiv:2608.06370v1 Announce Type: new Abstract: Tool use transforms LLMs into agents that act beyond their training data, and for code-capable models,…
面向软件工程智能体的技能多目标优化新方法,值得一读
arXiv:2604.09297v3 Announce Type: replace-cross Abstract: Agent skills are increasingly used to configure coding agents for software engineering (SE) …
多模态大模型能否胜任CEO决策?这篇论文用实验揭示视觉信息与决策能力之间的鸿沟,视角新颖。
arXiv:2608.05864v1 Announce Type: new Abstract: Large language models are increasingly applied as autonomous decision-making agents. However, in execu…
跳出循环的隐式监督,多保真贝叶斯优化迎来新范式,降本增效直击采样痛点。
arXiv:2608.04113v1 Announce Type: cross Abstract: Black-box optimization is a ubiquitous problem in science and engineering, often dealing with expens…
探究大模型能否跳出单一推理路径,自主发现多元异构推断,视角新颖。
arXiv:2608.02867v1 Announce Type: new Abstract: Although reinforcement learning with verifiable rewards (RLVR) has improved the performance of large l…