1
Stateful Guardrails for Multi-Turn LLM Systems: A Conversational Risk Accumulation Framework
多轮LLM对话安全的新框架,通过状态跟踪实现风险累积防护,应对长对话中的隐性威胁。
arXiv:2607.19361v1 Announce Type: new Abstract: Most safety guardrails for large language models (LLMs) evaluate each prompt-response pair in isolatio…