AgentDish directory

sandboxing

Accepted listings with this tag.

Listing Category Score Trend Checked
#35 ↓ -2
BoundaryBench

BoundaryBench is an open-source benchmark for coding agents running under hardened sandbox policies. It compares harnesses like Claude Code, Codex, Terminus 2, and Grok on Terminal-Bench tasks and includes quickstart, policy controls, and result export tools.

Developer Tools / Code Assistant 90 ↓ -2 26 days ago Details
#59 ↓ -2
Omnigent

Open-source AI agent framework and meta-harness for orchestrating Claude Code, Codex, Cursor, Pi, and custom agents with policies, sandboxing, and cross-device collaboration.

AI Development / Agent frameworks 90 ↓ -2 68 days ago Details
#220 ↓ -3
Termic

Desktop app for running CLI coding agents like Claude Code and Codex in parallel git worktrees, with built-in diff review, multi-repo support, broadcast, and sandboxing.

Developer Tools / Code Assistant 88 ↓ -3 37 days ago Details
#333 ↓ -4
XSAF

XSAF is a small, fetch-native TypeScript framework for building secure AI agents. The page shows its core API, agent lifecycle, tool pipeline, MCP support, session handling, and offline testing story.

Developer Tools / Code Assistant 87 ↓ -4 25 days ago Details
#457 ↑ +2
Runeward

Open-source governance harness for AI agents with policy enforcement, human approvals, isolated execution, budgets, and signed audit evidence. It includes a CLI, REST API, web dashboard, SDKs, and adapters for common agent frameworks.

Developer Tools / AI Agent Infrastructure 86 ↑ +2 26 days ago Details
#693 ↓ -3
askhuman.app

A web utility for creating end-to-end encrypted share links for single-file HTML generated by agents. It stores ciphertext for seven days and lets browsers decrypt and render the page from the URL fragment in a sandbox.

Developer Tools / API / Web Utility 85 ↓ -3 73 days ago Details
#834 ↓ -6
Democr.ai

A self-hosted Python runtime for building agentic AI apps with server-driven UI, audit logging, RBAC, sandboxing, and pluggable model orchestration.

Developer Tool / AI Application Framework 84 ↓ -6 46 days ago Details

An open-source reference implementation for autonomous vulnerability discovery and remediation with Claude. It includes Claude Code skills for threat modeling, scanning, triage, patching, plus a harness for running a recon → find → verify → report → patch pipeline.

Security / AI Security 84 ↓ -6 88 days ago Details
#953 ↓ -6
terminal-guardian-mcp

A secure Model Context Protocol server that gives AI assistants controlled terminal access with risk analysis, sandboxing, logging, filesystem protection, and optional Docker and Git features.

Developer Tools / MCP Servers 84 ↓ -6 101 days ago Details
#1002 ↓ -4
VT Code

Open-source Rust coding agent with LLM-native code understanding, shell safety, and support for multiple LLM providers with automatic failover.

Developer Tools / Code Assistant 84 ↓ -4 118 days ago Details
#1047 ↓ -3
mcpvessel

A CLI for running untrusted MCP servers in sandboxed containers with deny-by-default egress, traffic watching, and host approval workflows. It can also compose servers into agents and expose them over OCI registries.

Developer Tools / Security / Sandboxing 83 ↓ -3 27 days ago Details
#1357 ↑ +2
Dabs

Open-source CLI for spinning up disposable local or remote sandboxes for AI agents. It uses a manifest and Dockerfile to build fresh boxes, then exposes a single shell tool through MCP so an agent can work inside the sandbox without touching the host.

Developer Tools / AI Agent Infrastructure 81 ↑ +2 60 days ago Details
#1414 ↓ -1
venv-manager

A Go-based CLI for managing Python virtual environments, with MCP tools for AI agents, typed JSON snapshots, ephemeral execution, and file-watcher sync for evolving scripts.

Developer Tools / Python Development 80 ↓ -1 43 days ago Details

A Claude Code skill that launches a second model in a herdr split pane to adversarially review a diff, then checks each finding before returning a verdict.

Developer Tools / Code Review 78 ↑ +6 41 days ago Details
#1602 ↑ +6
Teleport-Env

Teleport-Env is an ultra-fast OS-level snapshot and rollback sandbox for autonomous coding agents, built with overlayfs and CRIU. It targets destructive agent testing, MCTS search loops, and reinforcement learning workflows that need rapid environment recovery.

Developer Tools / AI Agent Infrastructure 78 ↑ +6 96 days ago Details

Docker blog post about a real AI coding agent failure and how Docker Sandboxes aim to contain destructive execution mistakes.

Developer Tools / Code Assistant 75 → 0 92 days ago Details