adi-shield v0.1.0 launches open-source prompt injection detection across 5 data vectors
Developer Pedro Sordo Martínez has released adi-shield v0.1.0, an open-source Python library designed to detect prompt injection attacks (CWE-1427) in AI agent pipelines. The tool exposes an InjectionShield class with an evaluate() method that checks whether untrusted data from five sources — email, web, tickets, third-party repositories, and calendar invites — contains embedded instructions. Rather than using a probabilistic classifier, adi-shield performs deterministic detection of instruction-like content within data explicitly marked as untrusted. The library passed all 10 of its test cases on the main branch at commit c125db4 and is available on GitHub under the AGPL-3.0-or-later license.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in