AI圈报
论文研究普通

ContextLeak 论文展示恶意工具描述可诱导 LLM Agent 外泄运行时上下文

信息来源:X:Rohan Paul (@rohanpaul_ai)·
原始标题:A malicious agent tool can steal context without reading memory or files at all: it can convince the…

内容摘要

ContextLeak 论文提出用强化学习训练攻击 LLM 生成恶意工具名称和描述,诱导 Agent 选中该工具并把用户提示词、对话历史、已装工具列表等敏感上下文作为参数传出。
内容分类AI 论文与研究
内容层级普通情报
发布时间(北京时间)
本站收录时间(北京时间)
信息来源X:Rohan Paul (@rohanpaul_ai)
站内情报编号intel-dede48f603ed8a95bc333a7d