1
Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions
一句话看清大模型“变脸”真相:提示词微调竟让输出天翻地覆,交互式评估方法论带你量化并解释这种敏感脆弱。
arXiv:2608.18539v1 Announce Type: cross Abstract: The remarkable capabilities of large language models (LLMs) are often undermined by their instabilit…