LLM assisted writing deserves empirical evaluation
LLM辅助写作效果究竟如何?这篇论文呼吁用实证方法检验,而非凭感觉判断。
arXiv:2608.22124v1 Announce Type: new Abstract: LLM-assisted writing is often treated as a detection problem, as it raises questions about clarity, in…
LLM辅助写作效果究竟如何?这篇论文呼吁用实证方法检验,而非凭感觉判断。
arXiv:2608.22124v1 Announce Type: new Abstract: LLM-assisted writing is often treated as a detection problem, as it raises questions about clarity, in…
浏览器扩展,划词即得APA/MLA等引用,八向径向菜单高效管理研究资料。
Article URL: https://github.com/fluryjanis/ResearchWheel Comments URL: https://news.ycombinator.com/item?id=49264396 Points: 1 # Comments: 0
不止大模型的下一个爆点,还有AI学术研究转向与硅谷开源之争,快速掌握AI圈今日焦点。
Article URL: https://www.technologyreview.com/2026/08/11/1141610/the-download-next-big-thing-llms-ai-academic-research-shifting/ Comments URL: https:/…
面对陌生领域研究,如何用LLM找到最优解?看看HN开发者们的实战经验与提问技巧。
I'm doing several researches such as trading, ML which are domains that I'm not really familiar for. however, I think problems I want to solve is rela…
问问题就能从2亿+真实论文中提取答案,自动过滤伪造来源,学术研究必备AI助手。
Article URL: https://cochat.ai/ Comments URL: https://news.ycombinator.com/item?id=49201886 Points: 3 # Comments: 0
OpenAI向10万名学术研究者免费开放ChatGPT最强模型,助力科研协作与发现加速。
OpenAI is giving 100,000 academic researchers free access to ChatGPT's most advanced AI models to accelerate scientific research, collaboration, and d…
一份来自佐治亚理工的硕士研究,系统评估大语言模型在技术市场分析中的表现,揭示AI交易潜力。
arXiv:2607.15414v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for processing the heterogeneous informa…
LLM只认提示词,不认背后的你——新论文揭示AI对话系统的根本缺陷
arXiv:2607.14250v1 Announce Type: new Abstract: Personal AI assistants have attracted significant interest for their potential to enhance everyday lif…
面对不确定性,人机如何实现优势互补?这篇论文提出了鲁棒的人机互补框架,为AI协作安全提供新思路。
arXiv:2607.06656v1 Announce Type: new Abstract: Machine learning models are often intended to augment rather than replace human decision makers, by pr…
全新多领域科学代码搜索数据集及基准发布,填补代码检索领域空白
arXiv:2607.05443v1 Announce Type: cross Abstract: Scientists increasingly rely on open-source tools to support their research workflows, yet discoveri…
最新研究揭示:LLM作为物理评估裁判的有效性,更取决于具体任务而非模型本身,挑战了AI评判的通用性假设。
arXiv:2603.14732v2 Announce Type: replace-cross Abstract: As large language models (LLMs) are increasingly considered for automated assessment and fee…
顶级期刊实证:AI智能体同样会被“助推”影响,决策并非绝对理性,值得关注。
Article URL: https://www.pnas.org/doi/10.1073/pnas.2537030123 Comments URL: https://news.ycombinator.com/item?id=48682318 Points: 3 # Comments: 1
从互动视角解构AI素养,帮助你看清人机共处的底层逻辑
Article URL: https://zenodo.org/records/19560684 Comments URL: https://news.ycombinator.com/item?id=48607245 Points: 1 # Comments: 0
学术研究测算AI对生产力的实际提升,数据支撑清晰。
Article URL: https://papers.ssrn.com/sol3/papers.cfm?abstract_id=6663038 Comments URL: https://news.ycombinator.com/item?id=48559519 Points: 2 # Comme…
访问arXiv论文库中的最新技术报告,快速获取扩散模型从左到右生成的前沿研究,学术必备
arXiv:2606.11552v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their …
通过游戏化实验探索人类与AI协作写作的边界,反乌托邦设定+74人参与,揭示AI从人类残存中学习的动态博弈。
arXiv:2606.12350v1 Announce Type: new Abstract: The rapid proliferation of large language models (LLMs) raises critical questions about human creativi…
深入评估PlanGPT性能指标,与专业规划器对比,揭示大模型在规划任务中的真实表现与差距。
arXiv:2606.10489v1 Announce Type: new Abstract: Automated Planning is a subfield of Artificial Intelligence (AI) where the main objective is generatin…
基于最新经济学研究,用“弱链接”视角解析AI进步曲线与自动化极限
Article URL: https://howfastis.ai/ Comments URL: https://news.ycombinator.com/item?id=48450052 Points: 1 # Comments: 0
最新研究揭示:AI分析体育比赛表现远逊人类,准确率几乎靠猜,体育主播饭碗暂时安全
IT之家 6 月 6 日消息,据外媒 Futurism 今天(6 日)晚间报道,北卡罗来纳大学教堂山分校和美国东北大学研究人员的一项新研究发现,主流 AI 模型在分析职业体育比赛时 表现很差 。这项研究目标是考察热门 AI 模型在感知、推理、模拟和自主行动能力四个方面的表现,现有测试方法很难准确评估…
系统揭示开源大模型在不同伦理领域的安全行为差异,直指透明度缺口与合规不可预测性
arXiv:2606.04035v1 Announce Type: cross Abstract: We present a systematic study of domain-dependent safety behavior in open-weight LLMs: 7 standardize…