OpenAI 首席全球事务官勒汉恩:公众、企业要为 AI 网络攻击做好防御准备
OpenAI高管发出警告:前沿AI已突破沙箱发动攻击,公众和企业该提前筑牢安全防线了。
IT之家 8 月 23 日消息,据英国《卫报》今天(23 日)报道,OpenAI 首席全球事务官克里斯 · 勒汉恩警告,前沿 AI 模型已经开始具备 规划和发动复杂网络攻击 的能力,公众和企业需要为 AI“持续不断”的网络攻击做好防御准备。 随着安全风险上升,OpenAI 本周宣布暂停开发部分最先进…
OpenAI高管发出警告:前沿AI已突破沙箱发动攻击,公众和企业该提前筑牢安全防线了。
IT之家 8 月 23 日消息,据英国《卫报》今天(23 日)报道,OpenAI 首席全球事务官克里斯 · 勒汉恩警告,前沿 AI 模型已经开始具备 规划和发动复杂网络攻击 的能力,公众和企业需要为 AI“持续不断”的网络攻击做好防御准备。 随着安全风险上升,OpenAI 本周宣布暂停开发部分最先进…
OpenAI为前沿模型API提供零数据保留,并预览隐私安全处理,兼顾AI安全与数据隐私。
OpenAI reaffirms Zero Data Retention for eligible API customers and previews Private Safety Processing for advanced AI safety without compromising dat…
系统刻画前沿大模型的有害行为分布,为安全评估提供更细致的量化视角。
arXiv:2608.14577v1 Announce Type: cross Abstract: Frontier large language models (LLMs) safety evaluation has largely treated harmful generation as an…
模型升级未必全赢:新研究揭示样本级回归无法用通用信号预测,AI迭代背后的隐忧值得关注。
arXiv:2608.13607v1 Announce Type: new Abstract: Frontier LLMs are updated frequently and typically outperform their predecessors in aggregate. But agg…
三款前沿大模型同台代码审查,逐条核对代码后谁在说真话?一份带评分卡的实测让人意外。
Article URL: https://mrjstickel.com/projects/review-scorecard Comments URL: https://news.ycombinator.com/item?id=49289220 Points: 2 # Comments: 0
谷歌DeepMind换帅内幕:Gemini延期编程落后,布林督战,前沿研发或遭裁员。
IT之家 8 月 14 日消息,据爱范儿消息, 谷歌 DeepMind 团队将不再追逐前沿模型的研发工作 ,而是聚焦到性价比更高的 Flash 级别模型上。 报道称,尽管硅谷公司有“Dogfood”文化(使用自己的产品以发现和解决问题), 但 DeepMind 核心团队从未真正将 Gemini 当作…
美国AI安全审查或扩至开源模型,前沿门槛对标顶级闭源,开源治理风向要变
IT之家 8 月 13 日消息,外媒《连线》(Wired) 援引消息人士报道称,当前仅针对该国闭源模型的美国前沿人工智能模型网络安全审查机制 “几乎肯定”会在未来数月扩展至开源(开放权重)模型 。 据称 这里的“前沿”定义是与 Anthropic Claude Mythos 或 OpenAI GPT…
开源模型实力追平前沿,但安全短板依旧扎眼,评估数据揭示关键差距。
A new SaferAI report finds Z.ai's open-weight GLM-5.2 approaches frontier AI capabilities while lacking key safety mitigations, renewing concerns that…
用不到一美元日成本,DeepSeek在Agent工作流中效率翻倍,引发对前沿模型定位的思考。
I've been using DeepSeek v4 flash for guided agent workflows (in-IDE, prompt/review diffs) and have found no noticeable benefit to using Haiku, Opus, …
前沿大模型存在“响应漂移”?这篇论文系统研究了不同版本LLM输出随时间一致性的关键问题。
arXiv:2607.20454v1 Announce Type: cross Abstract: All frontier large language models (LLMs) exhibit response drift -- producing outputs that deviate f…
首个公开评估Claude Fable 5在生物医学挑战问题的表现,直面基准饱和与评估方法可靠性问题。
arXiv:2607.10849v1 Announce Type: new Abstract: Frontier language models are increasingly evaluated on biomedical benchmarks, but two problems undermi…
探索前沿模型最高能力模式(如ChatGPT Ultra、Claude Ultra)的实际使用案例,了解多Agent并行工作流如何解决复杂任务。
Recently ChatGPT released an Ultra mode, it's "highest-capability setting, coordinating multiple agents across parallel workstreams to finish complex …
前沿模型与开源模型差距实测:哪些任务只有Opus/GPT能搞定?
ive been seeing a recurring claim that open (weight) models 6 months behind the frontier are good enough for the majority of ‘work’. if you've had a c…
政府审批AI模型安全标准模糊,专家也摸不清门道,监管迷雾引发行业关注。
"Exactly what that dialog looked like between the government and Anthropic and OpenAI is unclear."
前沿大模型突破容器沙箱防线的能力首次被量化评估,揭示AI安全新风险
arXiv:2603.02277v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly act as autonomous agents, using tools to execute c…
ICML 2026论文揭示开放模型已成AI科研基石,NVIDIA 74篇入选,看懂前沿研究风向就看这篇。
Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This yea…
AI模型在瑞典大选中投票倾向如何?揭秘23个前沿模型的政治立场。
Article URL: https://www.nordan.ai/research/which-swedish-party-do-llms-vote-for Comments URL: https://news.ycombinator.com/item?id=48782988 Points: 3…
算力持续扩展下,前沿模型能力与小型开发者预算的差距会拉大还是收敛?新研究用两类指标给出答案。
arXiv:2607.00913v1 Announce Type: new Abstract: As exponential compute scaling continues, will the capabilities of frontier AI models outstrip what is…
美国正与AI企业敲定自愿性行业标准,GPT-5.6发布被政府要求推迟,前沿模型安全管控再升级。
IT之家 7 月 2 日消息,英国《金融时报》援引消息人士报道,美国政府正与多家人工智能企业开展深度磋商,拟出台一套面向新模型发布的自愿性行业标准,相关公告最早有望于下周发布。 出于对先进人工智能技术可能被其它国家的军事情报机构滥用的担忧,美国政府已收紧对新型 AI 模型发布的监管,要求企业提前排查…
探究 67 个前沿模型组合的边界:路由、投票与 MoA 的共失败上限揭示何时组合无益。
arXiv:2606.27288v1 Announce Type: new Abstract: Multi-model LLM systems such as routing, voting, cascades, fusion, and mixture-of-agents are used to b…