华为云联合瑞金医院发布瑞智病理大模型 RuiPath 2.0:7B 参数,覆盖 19 个常见癌种
IT之家 8 月 30 日消息,瑞金医院今日联合华为云正式发布瑞智病理大模型 RuiPath 2.0,带来性能、诊断能力、病理诊断可解释性、罕见病种数据训练、轻量化部署五大升级。 据介绍, 本次发布的 RuiPath 2.0 模型参数规模为 7B ,覆盖 19 个常见癌种,诊断任务增至 205 项,…
IT之家 8 月 30 日消息,瑞金医院今日联合华为云正式发布瑞智病理大模型 RuiPath 2.0,带来性能、诊断能力、病理诊断可解释性、罕见病种数据训练、轻量化部署五大升级。 据介绍, 本次发布的 RuiPath 2.0 模型参数规模为 7B ,覆盖 19 个常见癌种,诊断任务增至 205 项,…
IT之家 8 月 30 日消息,据白鲸实验室消息,字节跳动原计划于 8 月推出的豆包大模型 2.2 将推迟面世。 多位接近字节的人士称,延期的原因是,字节内部希望通过更充分的预训练和后训练,把编程、工具调用和 Agent 能力拉上去。 用更长的研发周期换一次更明显的能力提升 。 Seed 现在同时面…
IT之家 8 月 29 日消息,据“国光量子”公众号 27 日推文,国光量超人工智能研究团队将相关量子技术引入大语言模型的决策与推理过程,并面向科研和学术领域,推出玄幂(Xenomi)· 大模型家族。玄幂在科研科普、路由和智能体决策等特定任务中,相较于主流开源模型展现出更稳定的专业理解、更清晰的任务…
西门子Xcelerator与普通软件货架最本质的区别。货架解决的是「把产品卖出去」,西门子Xcelerator要解决的是「让产品在真实工业场景中持续生长」。
用4MB记忆让冻结的大模型记住新知识,极致压缩且效果惊艳。
Article URL: https://github.com/riscoss63/sillage Comments URL: https://news.ycombinator.com/item?id=49439609 Points: 1 # Comments: 0
OpenAI CFO亲述全栈AI战略:芯片、算力、模型、产品四层协同,看懂智能如何变得更便宜、更强大。
OpenAI CFO Sarah Friar explains how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale…
阿里千问明晚开源基于Qwen4架构的多模态MoE模型Qwen3.8-Flash-Next,新一代架构值得期待。
IT之家 8 月 25 日消息,魔搭社区(ModelScope)显示,阿里千问 Qwen3.8-Flash-Next 模型将于北京时间 08 月 26 日 23:00 开源。 页面显示,阿里千问将发布 Qwen3.8-Flash-Next 和 Qwen3.8-Flash-Next-FP8 两个版本模…
OpenAI新模型Bel曝光,超10万亿参数直指AGI,AI竞赛再升级。
IT之家 8 月 26 日消息,消息源 @synthwavedd 今天(8 月 26 日)在 X 平台发布推文,爆料称 OpenAI 已完成预训练下一代模型, 内部代号为 Bel,是 Doug 模型的直接继任者,总参数量超过 10T。 消息称 OpenAI 希望将 Bel 模型打造成为 GPT-6 …
系统梳理大模型时代人机协作前沿进展,洞见AI增强人类智慧的路径与挑战。
arXiv:2403.04931v4 Announce Type: replace Abstract: As the capabilities of artificial intelligence (AI) continue to expand rapidly, Human-AI (HAI) Col…
用证据门控从结构上杜绝漏洞幻觉的AI渗透测试智能体,全离线运行,安全报告终于不是AI编的故事了
If you point any LLM at a target and ask it to "write a security report," it will confidently invent findings that aren't there: an imagined TLS weakn…
用强化学习让大模型在终端环境中学会通用行为,打开日常自动化新可能。
arXiv:2608.22631v1 Announce Type: new Abstract: Terminal agents are a compelling application of large language models (LLMs), with the potential to in…
破解AI看病难题:全新基准MedReaMM如何考核多模态大模型的专业诊断整合能力
arXiv:2608.22323v1 Announce Type: new Abstract: The application of Large Language Models (LLMs) to diagnostic decision-making has garnered growing int…
首个聚焦音视频大模型安全性的基准,曝光跨模态越狱漏洞,为多模态安全研究提供关键测试标准。
arXiv:2508.07173v3 Announce Type: replace Abstract: Omni-modal Large Language Models (OLLMs) that integrate visual, auditory, and textual processing f…
扩散语言模型的自蒸馏新方法登上ACL 2026,生成效率与质量或迎来新突破。
arXiv:2608.22898v1 Announce Type: new Abstract: Diffusion language models (DLMs) alleviate the inherent latency bottleneck of autoregressive (AR) larg…
不只是拒绝请求,更是防住有害动作:这项研究揭示安全训练在智能体环境中能否真正“扛住”优化冲刷。
arXiv:2603.02229v2 Announce Type: replace Abstract: Safety post-training has been studied extensively in single-step "chat" settings where safety typi…
十四种后处理都扛不住20条样本复学攻击?这项研究用边界校准让LLM遗忘跨过悬崖、真正稳固。
arXiv:2607.27836v2 Announce Type: replace Abstract: Large language model unlearning is consistently fragile under relearn attacks. On TOFU, fine-tunin…
集合LoRA适配器构建置信区间,让大模型区分“不知道”与“不确定”,告别过度自信的胡说八道。
arXiv:2608.23244v1 Announce Type: cross Abstract: Large language models (LLMs) often produce fluent but incorrect answers with unwarranted confidence.…
分子科学遇上大模型智能体,从架构设计到自主科研,一文看尽前沿突破。
arXiv:2608.23104v1 Announce Type: cross Abstract: Molecular science represents an important frontier for LLM-based agents. Unlike general agents that …
用粤语语法资源当试金石,实测大模型做知识驱动语法工程的成色,实验设计严谨
arXiv:2608.23448v1 Announce Type: new Abstract: This paper presents new Cantonese ParGram resources and evaluates LLMs for knowledge-driven grammar en…