Developer builds AgentGuard CI tool after testing 12 attack payloads on own repo
A developer created AgentGuard, an open-source CI security tool designed to detect malicious instructions hidden inside AI coding agent configuration files such as AGENTS.md, MCP server definitions, and hooks. To validate the tool, the author planted 12 distinct attack payloads in their own repository, including hidden instruction overrides, zero-width characters, typosquatted server hosts, and committed API keys. AgentGuard detected all 12 threats deterministically without using an LLM in the scan path, and produced zero false positives on a clean control repository. A scan of 30 public repositories on GitHub, including those from Google, Microsoft, and Nextcloud, uncovered five with critical findings. The tool is MIT-licensed, free for public repositories on the GitHub Marketplace, and its own CI gate blocked the developer's pull request during a self-test, confirming it works as designed.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in