Skip to content
#

agent-security-tools

Here are 10 public repositories matching this topic...

Whitebox & Blackbox AI red-teaming framework for LLMs & Agentic AI apps. It analyzes your app's source code to discover tools, roles, and guardrails, then generates new attacks chains across several categories and adapts over multiple multi turn rounds to find vulnerabilities

  • Updated Sep 11, 2026
  • Python

🛡️ Open-source AI security scanner & LLM red-teaming platform. Test LLM APIs, chatbots, agents, MCP servers & RAG for prompt injection, jailbreaks, data leaks & unsafe tool use — with OWASP LLM Top 10 mapping and plain-English, audit-ready reports.

  • Updated Jul 9, 2026
  • Python

Manual Verification Protection + AI Selector Compatible with OpenClaw . This is a complete TypeScript plugin project—a true implementation of the `before_tool_call` hook that can intercept tool calls, follow OpenClaw’s built-in approval workflow, and features an approval window mechanism.

  • Updated Aug 26, 2026
  • TypeScript

A free course on AI agent security: prompt injection, tool poisoning, memory attacks, MCP supply chain, CaMeL, information-flow control, sandboxing and red-teaming. 27 chapters, runnable Python, interactive labs, six projects.

  • Updated Sep 14, 2026
  • JavaScript

Runtime capability governance for AI agents. Kingpin separates what an agent notices or believes from what it is authorized to do, using scoped leases, human review, retry controls, prompt-injection handling, and deterministic audit behavior.

  • Updated Sep 14, 2026
  • TypeScript

Daedalab is a control layer for AI agents that intercepts tool calls, blocks dangerous actions, enforces spending limits, and keeps a complete audit trail, giving users control over what their agents can do before it happens.

  • Updated Sep 8, 2026
  • TypeScript

Add this topic to your repo

To associate your repository with the agent-security-tools topic, visit your repo's landing page and select "manage topics."

Learn more