Why I still don’t use Claude (and why a “cheap model” is enough for me)
工程师亲身验证:高价AI并非必需,可预测性才是关键
In most discussions today, it feels like using advanced AI models has become a status signal: higher usage, bigger bills, “Claude maxed out again”, et…
工程师亲身验证:高价AI并非必需,可预测性才是关键
In most discussions today, it feels like using advanced AI models has become a status signal: higher usage, bigger bills, “Claude maxed out again”, et…
提出4/δ界限,为LLM验证器系统提供形式化保证,破解可预测性难题,值得关注。
arXiv:2512.02080v3 Announce Type: replace Abstract: The integration of Formal Verification tools with Large Language Models (LLMs) offers a path to sc…
差分隐私的粗粒度保护之外,这项研究提出用可预测性做更精细的隐私度量,值得关注。
arXiv:2606.20546v1 Announce Type: new Abstract: Differential privacy (DP) ensures rigorous individual-level privacy guarantees against even the most k…
LLM检索让广告推荐更稳定可预测,看论文如何用大模型优化推荐系统
arXiv:2605.21969v1 Announce Type: cross Abstract: Traditional ads recommendation systems have primarily focused on optimizing for prediction accuracy …
研究发现LLM的幻觉并非随机,而是与模型规模和主题频率呈可预测的比例关系。
arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked fact…