苹果部分语言版本 M5 Ultra Mac Studio 新闻稿提及基于 PCIe Gen6 的固态硬盘架构
IT之家 8 月 26 日消息,Apple(苹果)昨日推出了搭载 M5 Ultra 的 Mac Studio 紧凑型工作站。值得注意的是, 仅有部分语言版本的相关新闻稿提及了基于 PCIe Gen6 的固态硬盘架构 。 简体中文(有) 基于 PCIe 第六代架构的新一代固态硬盘架构使存储性能提升最高…
IT之家 8 月 26 日消息,Apple(苹果)昨日推出了搭载 M5 Ultra 的 Mac Studio 紧凑型工作站。值得注意的是, 仅有部分语言版本的相关新闻稿提及了基于 PCIe Gen6 的固态硬盘架构 。 简体中文(有) 基于 PCIe 第六代架构的新一代固态硬盘架构使存储性能提升最高…
IT之家 8 月 26 日消息,手办厂商 Good Smile Company(GSC)宣布,将为语言学习平台多邻国(Duolingo)的热门吉祥物“多儿(Duo)”推出为黏土人模型。目前官方尚未公布具体售价及发售时间,仅表示后续将公开更多信息。 只要使用过免费语言学习 App 多邻国的用户,想必都…
让AI在潜空间里“默想”而非逐字推演,这项研究颠覆了传统思维链范式
arXiv:2412.06769v4 Announce Type: replace Abstract: Large language models (LLMs) are typically constrained to reason in the language space, where they…
扩散语言模型的自蒸馏新方法登上ACL 2026,生成效率与质量或迎来新突破。
arXiv:2608.22898v1 Announce Type: new Abstract: Diffusion language models (DLMs) alleviate the inherent latency bottleneck of autoregressive (AR) larg…
面向全科疾病诊断的ClinicalGPT-R1,专注提升大模型临床推理能力,医学AI新突破。
arXiv:2504.09421v3 Announce Type: replace-cross Abstract: Recent advances in reasoning with large language models (LLMs)has shown remarkable reasoning…
用粤语语法资源当试金石,实测大模型做知识驱动语法工程的成色,实验设计严谨
arXiv:2608.23448v1 Announce Type: new Abstract: This paper presents new Cantonese ParGram resources and evaluates LLMs for knowledge-driven grammar en…
揭秘LLM路由差距的根源:任务类型比模型选择更关键,实测14模型两次运行结果并不完全一致。
arXiv:2608.23023v1 Announce Type: new Abstract: An LLM router picks which model should answer each query. The appeal is that models fail on different …
精神病学专用模型对战通用聊天机器人,评测医疗问答的临床准确性,AI落地医疗的新探索。
arXiv:2608.22797v1 Announce Type: new Abstract: Background This study was designed to evaluate whether a domain-specific large language model (LLM) tr…
精准识别生物医学论文中的AI辅助写作痕迹,为学术诚信提供数据支撑,多引擎检测更可靠
IT之家 8 月 25 日消息,据外媒 phys 报道,近期 Lena Holzwarth 带头的研究团队分析了来自 PubMed Central 平台的 119 万余篇英文生物医学论文,发现生成式 AI 正迅速融入学术论文写作。 研究人员通过分析论文中的“AI 高频用词”(例“delves”“ex…
揭示低资源语言在向量空间中的几何结构,为多语言模型优化提供全新视角。
arXiv:2608.23358v1 Announce Type: new Abstract: The performance gap between low- and high-resource languages in LLMs is widely known, but it remains u…
从星系到语言模型,AstroPT揭示了预训练模型如何习得结构知识,为LLM研究提供跨学科洞见。
arXiv:2608.22614v1 Announce Type: new Abstract: Interpretability research increasingly asks when concepts emerge during training and whether linear pr…
用强化学习让大模型学会「讲道理」,逆合成预测从黑箱猜测走向可解释推理——AI驱动分子设计又进一步。
arXiv:2507.17448v2 Announce Type: replace-cross Abstract: Retrosynthetic planning is a cornerstone of organic synthesis and drug discovery. Yet existi…
揭秘架构如何成为编码智能体的能力均衡器,为提升代码生成质量提供新视角。
arXiv:2608.21747v1 Announce Type: cross Abstract: LLM-based coding agents generate complete software systems from high-level descriptions, yet little …
扩散语言模型推理提速新思路,用收敛感知机制减少计算浪费,值得关注。
arXiv:2608.22646v1 Announce Type: new Abstract: Diffusion language models can generate many tokens in parallel, but they still require repeated denois…
让大模型按需调整推理深度,AdaR框架为复杂问题动态分配计算资源,兼顾效率与准确。
arXiv:2510.04617v3 Announce Type: replace Abstract: Mathematical reasoning is a primary indicator of large language models (LLMs) intelligence. Howeve…
小孩学语言只需少量样本,大模型却要吞下整座城市几代人的语料,这个谜题正动摇AI的底层逻辑。
People have been talking to each other for at least 100,000 years, as best we can tell. And in all that time, there has been only one thing in the wor…
零样本大模型能否替代精调NLU?这项研究用ATIS和CLINC150实测,给出基于意图空间的决策框架,生产环境选型必看。
arXiv:2608.20371v1 Announce Type: cross Abstract: A common claim is that zero-shot large language models (LLMs) can replace fine-tuned NLU classifiers…
拆解LLM心理治疗每步动作,精准测量并引导对话走向,让AI咨询更可控。
arXiv:2608.21325v1 Announce Type: new Abstract: Users increasingly turn to large language models for emotional support, yet little is known about how …
深入探究AI意识之谜,看科学家如何在LLM中寻找思维的火花,颠覆你对智能的认知。
Article URL: https://www.economist.com/interactive/briefing/2026/08/20/the-search-for-consciousness-inside-llms Comments URL: https://news.ycombinator…
不量化不租GPU,用C语言流式加载权重,让284B大模型在3.2GB内存跑起来,揭秘极简推理的实现思路。
DeepSeek-V4-Flash has 284 billion parameters and takes up about 160GB on disk. My laptop does not have 160GB of RAM. It doesn't even have 32GB. It ran…