AgentDish directory
sandboxing
Accepted listings with this tag.
| Listing | Category | Score | Trend | Checked | |
|---|---|---|---|---|---|
|
#35
↓ -2
BoundaryBench
BoundaryBench is an open-source benchmark for coding agents running under hardened sandbox policies. It compares harnesses like Claude Code, Codex, Terminus 2, and Grok on Terminal-Bench tasks and includes quickstart, policy controls, and result export tools. |
Developer Tools / Code Assistant | 90 | ↓ -2 | 26 days ago | Details |
|
#59
↓ -2
Omnigent
Open-source AI agent framework and meta-harness for orchestrating Claude Code, Codex, Cursor, Pi, and custom agents with policies, sandboxing, and cross-device collaboration. |
AI Development / Agent frameworks | 90 | ↓ -2 | 68 days ago | Details |
|
#220
↓ -3
Termic
Desktop app for running CLI coding agents like Claude Code and Codex in parallel git worktrees, with built-in diff review, multi-repo support, broadcast, and sandboxing. |
Developer Tools / Code Assistant | 88 | ↓ -3 | 37 days ago | Details |
|
#333
↓ -4
XSAF
XSAF is a small, fetch-native TypeScript framework for building secure AI agents. The page shows its core API, agent lifecycle, tool pipeline, MCP support, session handling, and offline testing story. |
Developer Tools / Code Assistant | 87 | ↓ -4 | 25 days ago | Details |
|
#457
↑ +2
Runeward
Open-source governance harness for AI agents with policy enforcement, human approvals, isolated execution, budgets, and signed audit evidence. It includes a CLI, REST API, web dashboard, SDKs, and adapters for common agent frameworks. |
Developer Tools / AI Agent Infrastructure | 86 | ↑ +2 | 26 days ago | Details |
|
#693
↓ -3
askhuman.app
A web utility for creating end-to-end encrypted share links for single-file HTML generated by agents. It stores ciphertext for seven days and lets browsers decrypt and render the page from the URL fragment in a sandbox. |
Developer Tools / API / Web Utility | 85 | ↓ -3 | 73 days ago | Details |
|
#834
↓ -6
Democr.ai
A self-hosted Python runtime for building agentic AI apps with server-driven UI, audit logging, RBAC, sandboxing, and pluggable model orchestration. |
Developer Tool / AI Application Framework | 84 | ↓ -6 | 46 days ago | Details |
|
#929
↓ -6
Defending Code Reference Harness
An open-source reference implementation for autonomous vulnerability discovery and remediation with Claude. It includes Claude Code skills for threat modeling, scanning, triage, patching, plus a harness for running a recon → find → verify → report → patch pipeline. |
Security / AI Security | 84 | ↓ -6 | 88 days ago | Details |
|
#953
↓ -6
terminal-guardian-mcp
A secure Model Context Protocol server that gives AI assistants controlled terminal access with risk analysis, sandboxing, logging, filesystem protection, and optional Docker and Git features. |
Developer Tools / MCP Servers | 84 | ↓ -6 | 101 days ago | Details |
|
#1002
↓ -4
VT Code
Open-source Rust coding agent with LLM-native code understanding, shell safety, and support for multiple LLM providers with automatic failover. |
Developer Tools / Code Assistant | 84 | ↓ -4 | 118 days ago | Details |
|
#1047
↓ -3
mcpvessel
A CLI for running untrusted MCP servers in sandboxed containers with deny-by-default egress, traffic watching, and host approval workflows. It can also compose servers into agents and expose them over OCI registries. |
Developer Tools / Security / Sandboxing | 83 | ↓ -3 | 27 days ago | Details |
|
#1357
↑ +2
Dabs
Open-source CLI for spinning up disposable local or remote sandboxes for AI agents. It uses a manifest and Dockerfile to build fresh boxes, then exposes a single shell tool through MCP so an agent can work inside the sandbox without touching the host. |
Developer Tools / AI Agent Infrastructure | 81 | ↑ +2 | 60 days ago | Details |
|
#1414
↓ -1
venv-manager
A Go-based CLI for managing Python virtual environments, with MCP tools for AI agents, typed JSON snapshots, ephemeral execution, and file-watcher sync for evolving scripts. |
Developer Tools / Python Development | 80 | ↓ -1 | 43 days ago | Details |
|
#1564
↑ +6
Herdr Adversarial Review
A Claude Code skill that launches a second model in a herdr split pane to adversarially review a diff, then checks each finding before returning a verdict. |
Developer Tools / Code Review | 78 | ↑ +6 | 41 days ago | Details |
|
#1602
↑ +6
Teleport-Env
Teleport-Env is an ultra-fast OS-level snapshot and rollback sandbox for autonomous coding agents, built with overlayfs and CRIU. It targets destructive agent testing, MCTS search loops, and reinforcement learning workflows that need rapid environment recovery. |
Developer Tools / AI Agent Infrastructure | 78 | ↑ +6 | 96 days ago | Details |
|
Docker blog post about a real AI coding agent failure and how Docker Sandboxes aim to contain destructive execution mistakes. |
Developer Tools / Code Assistant | 75 | → 0 | 92 days ago | Details |