Baikal: Structured Search for Deep Research over Data Lakes
面向数据湖的结构化搜索新框架,为深度研究提供高效精准的信息定位方案
arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of hete…
面向数据湖的结构化搜索新框架,为深度研究提供高效精准的信息定位方案
arXiv:2607.27726v1 Announce Type: cross Abstract: Deep research over data lakes requires an LLM agent to investigate evidence across thousands of hete…
可扩展的强化学习训练框架LiteResearcher问世,专为深度研究智能体打造,效率与性能双提升。
arXiv:2604.17931v3 Announce Type: replace Abstract: Reinforcement Learning (RL) has emerged as a powerful training paradigm for LLM-based agents. Howe…
大模型深度研究报告的逻辑质量如何量化?这项研究提出全新评估框架ReportLogic,直击AI报告的可信度痛点。
arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to syn…
深度研究任务中重新审视文本排序,挑战传统LLM+搜索API的固有范式。
arXiv:2602.21456v2 Announce Type: replace-cross Abstract: Deep research has emerged as an important task that aims to address hard queries that need e…
开源小插件让你在Claude Code里直接调用Google搜索、代码审查、深度研究和图像生成,无需离开编辑器。
Quick share for people who work with Claude Code, Codex or other agents. I built a Claude Code plugin that gives you agy access from inside CC without…
告别浅层搜索,Sakana AI的Marlin用8小时深度推理,自动生成100+页B2B商业研究报告
Tokyo-based AI startup Sakana AI has officially launched its first commercial product, Sakana Marlin . Billed as a " Virtual CSO " (Chief St…
用代理式大模型巧妙分解长周期任务,突破上下文窗口限制,实现深度研究新范式。
arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose…
从聊天记录自动构建知识源库,支持深度研究模式和Gemini 3.5模型,让研究更高效。
Google is making Gemini 3.5 the default model in NotebookLM
从AI包装器迈向自主进化代理,揭秘闭环学习架构实现深度研究与CI/CD自动化。
We are officially transitioning from the era of "AI wrappers" to the era of truly autonomous agentic systems. If you’ve spent any time building with L…
揭示深度研究代理中搜索时间污染导致基准性能虚高,为AI评估体系敲响警钟。
arXiv:2606.05241v1 Announce Type: cross Abstract: Public benchmarks enable fair and reproducible evaluation of LLM reasoning, but they become fragile …
探索AI自我进化的新范式:通过联合生成与评估实现深度研究的迭代提升
arXiv:2606.04507v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly adopted in daily applications, with deep resea…
针对深度研究代理面临的信息不确定与中间表示污染问题,提出可演化的心智模型显式调控方法。
arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing system…
谷歌Gemini智能家居版闹笑话,猫狗不分,袋鼠认成人,同时曝出多项新模型和订阅层级动态。
IT之家 5 月 26 日消息,澳大利亚网友 u/That_Car_Dude_Aus 昨日(5 月 25 日)在 Reddit 社区发帖, 表示智能家居版 Gemini 多次出现误判,把猫误认成浣熊,把袋鼠归类为“人”。 根据该网友反馈,Gemini for Home 在摄像头画面里频繁认错动物和车…
阿里千问全面升级,免费体验全新Qwen3.7-Max大模型,支持股票数据深度研究。
IT之家 5 月 22 日消息,千问 App 官方公众号宣布,千问 App、PC 端及网页端接入全新一代大模型 Qwen3.7-Max。 将千问 APP 更新至最新版(IT之家注:6.9.7 及以上)后,点击下方胶囊“Qwen3.7-Max”,或在 PC 端及网页端对话界面的“模型选择栏”中进行下拉…
提出Argus框架,用证据组装突破单轨迹限制,实现可扩展的深度研究代理。
arXiv:2605.16217v1 Announce Type: cross Abstract: Deep research agents have achieved remarkable progress on complex information seeking tasks. Even lo…