1
Before the Action: Benchmarking LLMs on Prospective Hypothesis Discovery
全新基准测试揭示大模型在科学假设发现前的推理能力,填补前瞻性评估空白。
arXiv:2607.15766v1 Announce Type: new Abstract: Large language models (LLMs) excel at answering pre-specified questions, yet their ability to navigate…