Open-source adversarial testing engine, SDK, and CLI for AI agents. Runs locally or against the Humanbound Platform.
-
Updated
Sep 9, 2026 - Python
Open-source adversarial testing engine, SDK, and CLI for AI agents. Runs locally or against the Humanbound Platform.
Multi-tier firewall for AI agents — blocks prompt injections, jailbreaks, and scope violations. Local tiers first; LLM judge only when uncertain.
Official GitHub Actions for Humanbound — adversarial security testing for AI agents in CI.
Two-stage security architecture to mitigate indirect prompt injection attacks in rich content via privilege separation
Adversarial research on multi-modal LM like LVLMs.
To associate your repository with the multimodal-security topic, visit your repo's landing page and select "manage topics."