Skip to content
arXiv cs.CL · Papers

AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents

arXiv:2607.26998v1 Announce Type: cross Abstract: Large language model (LLM) agents automate penetration testing through an observation-action loop, selecting actions based on observations returned by tools. This dependence allows defenders to inject deceptive observations that can mislead the agent's decision-making p