1
InferenceBench: A Benchmark for Open-Ended LLM Inference Optimization by AI Agents
首个针对AI代理优化开放式LLM推理的基准测试,填补评估空白。
arXiv:2607.20468v1 Announce Type: new Abstract: AI agents are increasingly used to automate research and development tasks, yet existing benchmarks ty…