AgentDish directory

agent security

Accepted listings with this tag.

Listing Category Score Trend Checked

An OWASP incubator project that protects AI agent memory from prompt injection, secret leakage, and tampering. It includes a Python library, policy-based controls, benchmarks, and integrations for agent frameworks like LangChain and AutoGen.

Developer Tools / AI Security 91 ↓ -3 103 days ago Details
#360 ↓ -4
bastiontrace

Bastiontrace is a Python library for analyzing AI agent tool-call traces to locate prompt injections, identify where they landed, and map the resulting blast radius.

Developer Tools / Security / AI Agent Monitoring 87 ↓ -4 17 hours ago Details
#605 ↑ +2
Lelu

Open-source authorization engine for AI agents that adds confidence-based gating, human review, policy-as-code, and audit logging. The repo shows quickstart code, local demo steps, SDK installs, and self-hosting options.

Developer Tool / AI Authorization / Security 86 ↑ +2 83 days ago Details
#727 ↓ -3
LLM Red Team Lab

A hands-on kit for authorized red teaming of locally run LLMs, with jailbreak techniques, prompt-injection scenarios, streaming attack visualization, and support for OpenAI-compatible local model servers.

Security / AI Red Teaming 85 ↓ -3 46 days ago Details
#889 ↓ -6
ModelFuzz

Open-source runtime guardrails for AI agents that intercept unsafe tool calls before they execute, aiming to block prompt-injection-driven exfiltration and other policy violations.

Developer Tools / AI Safety 84 ↓ -6 38 days ago Details
#907 ↓ -6
ASL V6

Open-source security research tool for auditing Python AI agents and codebases with AST analysis and Docker-based runtime verification.

Security / AI Security / Red Teaming 84 ↓ -6 46 days ago Details
#922 ↓ -6
Sunglasses

Open-source input scanner for AI agents that strips hidden instructions and scans text, files, images, PDFs, QR codes, audio, and video before they reach an agent.

Security / AI Security / Prompt Injection Defense 84 ↓ -6 50 days ago Details
#1175 ↓ -3
Lotor

Local-first approval and receipt logging for AI agent sessions, with tamper-evident records and signed approvals for consequential actions.

Developer Tools / AI Agent Tooling 83 ↓ -3 48 days ago Details
#1224 ↓ -3
Helm AI Kernel

A fail-closed execution firewall for AI agents that quarantines MCP tools, proxies OpenAI-compatible requests, and emits signed receipts for offline verification.

Developer Tools / AI Security 83 ↓ -3 92 days ago Details
#1349 ↓ -2
Vitrin OS

Open-source agent-first display server for safely running human and AI-driven GUI sessions with per-app isolation and capability-scoped authorization.

Developer Tools / AI Infrastructure 82 ↓ -2 47 days ago Details

An ICML 2026 research project page arguing that prompt injection comes from how LLMs misread roles, with an extended writeup, examples, and links to the paper, code, arXiv, and BibTeX.

Research / AI Safety 77 → 0 80 days ago Details
#1894 → 0
VAIBot

VAIBot is an AI governance product focused on agent security. This page explains its prompt-injection threat model and positions the product as a circuit breaker that gates tool calls and other egress actions before they execute.

AI Governance / Agent Security 75 → 0 63 days ago Details
#1913 ↓ -1
Oconee Runtime

Oconee Runtime is an enterprise AI governance product focused on policy enforcement for browser AI and coding agents. The page explains how it evaluates proposed agent actions against policy and supports allow, warn, and block outcomes.

Developer Tools / Code Assistant 74 ↓ -1 8 days ago Details

A Reco security research article showing an AI-powered agent that maps Salesforce Experience Cloud sites, probes exposed objects and Apex methods, and attempts autonomous exploitation to find data exposure.

AI Security / Agent Security 72 ↑ +1 92 days ago Details