每日 Harness 开源 · Source
返回本期 · Back to 2026-06-11

论文 · Papers2026-06-11 · Thursday, June 11, 2026

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

arxiv.org原文 ↗

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation
这篇综述把 LLM agent 安全从“输出安全”扩展到工具、权限、记忆和环境动作组成的系统问题。它梳理 247 篇论文,指出 prompt injection 与 tool-mediated control-flow hijacking 仍是主线,但 persistent state corruption 和 multi-agent propagation 正快速上升;它适合作为设计 agent 安全边界时的攻击面索引。
浏览

评论 · Comments