Benchmarking Trustworthiness of SLMs: Pre-trained vs. Compressed
预训练与压缩小模型谁更可信?IJCNN权威基准测试给出对比答案。
arXiv:2608.11981v1 Announce Type: new Abstract: Small Language Models (SLMs) have emerged as a more efficient alternative to traditional Large Languag…
预训练与压缩小模型谁更可信?IJCNN权威基准测试给出对比答案。
arXiv:2608.11981v1 Announce Type: new Abstract: Small Language Models (SLMs) have emerged as a more efficient alternative to traditional Large Languag…
首个为移动医疗场景下小语言模型(SLM)定制的标准化基准,填补健康监测评估空白
arXiv:2509.07260v5 Announce Type: replace Abstract: Mobile and wearable healthcare monitoring play a vital role in facilitating timely interventions, …
为LLM应用量身定制的轻量级安全护栏,用小模型实现高效防护。
arXiv:2607.18268v1 Announce Type: new Abstract: Real-world applications that use closed-source large language models (LLMs) need advanced safety measu…
急需低配本地模型?HN热帖推荐2B参数以下、内存<3GB的高效SLM方案
For my local project, I have been finding a good LLM or a SLM that consumes less than 2gb ram, can you get me any models that would aid me? Comments U…
EffGen让小语言模型通过生成示例实现自主推理,ICML 2026论文揭示高效智能体新路径。
arXiv:2602.00887v2 Announce Type: replace-cross Abstract: Most existing language model agentic systems today are built and optimized for large languag…
基于LLM到SLM的对抗性提示蒸馏,实现高效且隐蔽的越狱攻击新方法。
arXiv:2506.17231v3 Announce Type: replace Abstract: Current jailbreak attacks on large language models (LLMs) predominantly rely on LLMs themselves to…
学生眼中的AI反馈靠谱吗?这项研究对比了LLMs、SLMs与人类在技术写作反馈中的质量差异,为AI教育应用提供了实证参考。
arXiv:2601.11541v2 Announce Type: replace-cross Abstract: To address the scalability of feedback in computer science while mitigating the privacy and …
通过不确定性感知的机会传输与压缩,大幅降低混合语言模型中设备端与云端之间的通信开销。
arXiv:2505.11788v2 Announce Type: replace-cross Abstract: To support emerging language-based applications using dispersed and heterogeneous computing …
教你如何让小型语言模型学会判断何时该“求救”,避免盲目依赖昂贵大模型,提升Agent系统效率的突破性研究。
arXiv:2605.16604v1 Announce Type: new Abstract: Efficient agentic systems should incur expensive frontier-model costs only on decisions where a cheape…