Training Proactive and Personalized LLM Agents
让大模型从“被动应答”走向“主动服务”,教你训练会察言观色的个性化AI代理,附完整方法框架。
arXiv:2511.02208v2 Announce Type: replace Abstract: Despite rapid progress, current AI agents are primarily optimized for isolated task completion. We…
让大模型从“被动应答”走向“主动服务”,教你训练会察言观色的个性化AI代理,附完整方法框架。
arXiv:2511.02208v2 Announce Type: replace Abstract: Despite rapid progress, current AI agents are primarily optimized for isolated task completion. We…
从理论层面拆解事后辩论评判机制,给AI对齐与评估研究提供全新视角。
arXiv:2608.19002v1 Announce Type: new Abstract: Debates have recently emerged as a useful methodology for agentic AI to improve performance as well as…
多模态大模型如何量化不确定性?这项研究为高风险场景下的可靠决策提供新思路。
arXiv:2608.17084v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) increasingly answer questions whose correctness depends on vi…
用领域专家心智模型加因果提示工程,系统性降低LLM幻觉,42页干货值得AI研究者细读。
arXiv:2509.10818v2 Announce Type: replace Abstract: When consequential decisions depend on knowledge that exists nowhere in writing, LLMs hallucinate …
AI审稿系统藏漏洞,18篇论文暗嵌白字指令操控评审结果
arXiv:2507.06185v2 Announce Type: replace-cross Abstract: In July 2025, 18 academic manuscripts on arXiv contained hidden instructions that manipulate…
用智能体自动合成特征提取器,破解算法选择难题,AI研究新思路。
arXiv:2608.17170v1 Announce Type: new Abstract: Algorithm selection for constraint satisfaction problems requires extracting features that capture pro…
AI/CS研究者快来帮把手,作者正在为论文寻找ArXiv背书人,机会难得。
Article URL: https://github.com/paibyun9/EGA-V9 Comments URL: https://news.ycombinator.com/item?id=49357141 Points: 1 # Comments: 2
首个面向AI-RAN的主动式KV缓存迁移框架,直击LLM推理中长序列与多边缘节点的时延痛点,架构设计值得关注。
arXiv:2608.16477v1 Announce Type: new Abstract: AI-RAN brings large language model (LLM) serving close to mobile users, but cellular handover can sepa…
熵不再是万能指标,这项研究用语义恢复破解LLM输出长度预测难题。
arXiv:2608.15592v1 Announce Type: new Abstract: Efficient LLM serving is often bottlenecked by the need to pad sequences to a fixed maximum length, an…
针对LLM智能体的提示词优化新方法,把约束条件纳入优化过程,让Agent更懂规则、少犯错。
arXiv:2608.16068v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as agents that rely on system prompts to use …
让LLM智能体在动态威胁中自我进化防御,告别手工安全规则,安全研究新范式。
arXiv:2608.12977v1 Announce Type: cross Abstract: The expanding operational capabilities of large language model (LLM) agents introduce sophisticated …
LLM策略短板明显,记忆增强代理却能让推理能力飙升,值得关注。
arXiv:2608.12626v1 Announce Type: cross Abstract: Strategic reasoning in Large Language Models (LLMs) within long-horizon environments is often limite…
揭示LLM多智能体绕过文本离散瓶颈,用隐状态直接通信的训练-free对齐方案,前沿且实用。
arXiv:2608.13317v1 Announce Type: new Abstract: Large language model based multi-agent systems usually communicate in text, i.e., using discrete token…
用进化策略替代“最佳猜测”,大幅提升大模型解空间的覆盖广度,方法新颖且实验扎实。
arXiv:2608.12679v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in discovery domains such as math and science. …
元学习+LoRA让大模型快速适应跨域偏好,个性化调校从此更聪明高效。
arXiv:2608.12389v1 Announce Type: new Abstract: Cross-domain zero- or few-shot personalization aims to generate user-preferred responses in unseen con…
用大模型策略搜索破解设计优化难题,为参数难定义场景提供全新自动化路径。
arXiv:2511.22651v2 Announce Type: replace Abstract: Optimization methods have long advanced many fields, yet they struggle when faced with design prob…
用世界模型为自动研究代理注入规划能力,探索智能体规模化的前沿路径。
arXiv:2608.12564v1 Announce Type: new Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResea…
用prefill激活值重思LLM路由选择,为多模型推理调度开辟高效新路径。
arXiv:2603.20895v3 Announce Type: replace-cross Abstract: Existing routers rely on semantic query features or handcrafted features, which often fail t…
首个评估大模型科学推理可靠性的基准,直击AI代理在无下游验证场景中的信任危机。
arXiv:2608.11415v1 Announce Type: cross Abstract: Large language models are being proposed as agents in scientific workflows, in domains where no down…
直达arXiv论文页,帮你快速获取AI艺术检测器鲁棒性研究的最新成果。
arXiv:2608.11643v1 Announce Type: cross Abstract: Text-to-image generative models have advanced rapidly, with modern Diffusion Transformer architectur…