AgentDish directory
agent security
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#25
↓ -3
OWASP Agent Memory Guard
An OWASP incubator project that protects AI agent memory from prompt injection, secret leakage, and tampering. It includes a Python library, policy-based controls, benchmarks, and integrations for agent frameworks like LangChain and AutoGen. |
Developer Tools / AI Security | 91 | ↓ -3 | 103 days ago | Details |
|
#355
↓ -4
bastiontrace
Bastiontrace is a Python library for analyzing AI agent tool-call traces to locate prompt injections, identify where they landed, and map the resulting blast radius. |
Developer Tools / Security / AI Agent Monitoring | 87 | ↓ -4 | 7 hours ago | Details |
|
#599
↑ +2
Lelu
Open-source authorization engine for AI agents that adds confidence-based gating, human review, policy-as-code, and audit logging. The repo shows quickstart code, local demo steps, SDK installs, and self-hosting options. |
Developer Tool / AI Authorization / Security | 86 | ↑ +2 | 83 days ago | Details |
|
#721
↓ -3
LLM Red Team Lab
A hands-on kit for authorized red teaming of locally run LLMs, with jailbreak techniques, prompt-injection scenarios, streaming attack visualization, and support for OpenAI-compatible local model servers. |
Security / AI Red Teaming | 85 | ↓ -3 | 46 days ago | Details |
|
#877
↓ -6
ModelFuzz
Open-source runtime guardrails for AI agents that intercept unsafe tool calls before they execute, aiming to block prompt-injection-driven exfiltration and other policy violations. |
Developer Tools / AI Safety | 84 | ↓ -6 | 38 days ago | Details |
|
#895
↓ -6
ASL V6
Open-source security research tool for auditing Python AI agents and codebases with AST analysis and Docker-based runtime verification. |
Security / AI Security / Red Teaming | 84 | ↓ -6 | 46 days ago | Details |
|
#910
↓ -6
Sunglasses
Open-source input scanner for AI agents that strips hidden instructions and scans text, files, images, PDFs, QR codes, audio, and video before they reach an agent. |
Security / AI Security / Prompt Injection Defense | 84 | ↓ -6 | 50 days ago | Details |
|
#1163
↓ -3
Lotor
Local-first approval and receipt logging for AI agent sessions, with tamper-evident records and signed approvals for consequential actions. |
Developer Tools / AI Agent Tooling | 83 | ↓ -3 | 48 days ago | Details |
|
#1212
↓ -3
Helm AI Kernel
A fail-closed execution firewall for AI agents that quarantines MCP tools, proxies OpenAI-compatible requests, and emits signed receipts for offline verification. |
Developer Tools / AI Security | 83 | ↓ -3 | 92 days ago | Details |
|
#1336
↓ -2
Vitrin OS
Open-source agent-first display server for safely running human and AI-driven GUI sessions with per-app isolation and capability-scoped authorization. |
Developer Tools / AI Infrastructure | 82 | ↓ -2 | 47 days ago | Details |
|
#1802
→ 0
Prompt Injection as Role Confusion
An ICML 2026 research project page arguing that prompt injection comes from how LLMs misread roles, with an extended writeup, examples, and links to the paper, code, arXiv, and BibTeX. |
Research / AI Safety | 77 | → 0 | 80 days ago | Details |
|
#1878
→ 0
VAIBot
VAIBot is an AI governance product focused on agent security. This page explains its prompt-injection threat model and positions the product as a circuit breaker that gates tool calls and other egress actions before they execute. |
AI Governance / Agent Security | 75 | → 0 | 63 days ago | Details |
|
#1897
↓ -1
Oconee Runtime
Oconee Runtime is an enterprise AI governance product focused on policy enforcement for browser AI and coding agents. The page explains how it evaluates proposed agent actions against policy and supports allow, warn, and block outcomes. |
Developer Tools / Code Assistant | 74 | ↓ -1 | 8 days ago | Details |
|
#1998
↑ +1
Hacking Salesforce Sites With an LLM Agent
A Reco security research article showing an AI-powered agent that maps Salesforce Experience Cloud sites, probes exposed objects and Apex methods, and attempts autonomous exploitation to find data exposure. |
AI Security / Agent Security | 72 | ↑ +1 | 92 days ago | Details |