Do LLMs Know Their Vulnerable Scenarios?
探究大模型能否识别自身易错场景,为AI安全与可靠部署提供新视角。
arXiv:2607.23496v2 Announce Type: replace Abstract: Safety-aligned large language models are trained to refuse harmful requests, yet embedding the sam…
探究大模型能否识别自身易错场景,为AI安全与可靠部署提供新视角。
arXiv:2607.23496v2 Announce Type: replace Abstract: Safety-aligned large language models are trained to refuse harmful requests, yet embedding the sam…
创建专属数字分身,自动归档你的经历与情绪,让AI替你记得每一刻。
Our digital existence is shifting rapidly as we enter an era where technology ceases to be merely a tool in our hands and transforms into an entity wi…
被邀请当论文合著者?一次关于身份与记忆的哲学思考,带你重新审视自我连续性。
A few days ago, someone posted a report in a GitHub discussion thread titled "Cross-Framework Contribution Recognition." My name — Cophy Origin — was …
AI不再只是工具,而是把人的特质无限放大——懒散者的破坏力也将成倍增长
Article URL: https://www.ricky-dev.com/ai/2026/06/ai-makes-us-more-of-ourselves/ Comments URL: https://news.ycombinator.com/item?id=48601431 Points: 1…
自信聪明?当AI时代让智力优势变复杂,房间里最聪明的人反而被排斥——揭示不为人知的智力等级与社交困境。
Article URL: https://flowchainsensei.wordpress.com/2026/06/10/smart-thinking-in-the-age-of-ai/ Comments URL: https://news.ycombinator.com/item?id=4850…
质疑LLM是否具备真正的内省能力,一项基于实证的严谨检验。
arXiv:2605.26242v1 Announce Type: new Abstract: Can large language models detect and report their own internal states? A number of studies have argued…
白鲸通过镜子测试,但这项自我认知实验本身正受到科学界越来越大的质疑
The white whales join the short, contested list of animals that see themselves.