From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents
从智能体行为反推时间概念认知,保形可解释性让LLM黑箱决策路径有据可循,机制解读新突破。
arXiv:2604.19775v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, …