Security / AI Agent Security

Declaw Arena

An interactive CTF-style challenge where users try to break or exfiltrate data from sandboxed AI agents running in Declaw’s isolated runtime.

Clear27/30
Useful24/30
Specific17/20
Complete16/20
Declaw Arena screenshot

Why it was accepted

The page clearly presents a real AI-powered security challenge with multiple agent scenarios, visible difficulty levels, and concrete attack goals. It explains the product context, shows that the sandbox uses runtime policies and isolated sessions, and gives enough detail for a useful directory listing.

Weakness

The snapshot does not show technical docs, setup instructions, or how the arena maps to Declaw’s broader runtime product. It also doesn’t explain whether results are persisted, how challenges are scored, or whether the scenarios are meant purely for demos, training, or ongoing testing.

Review status

59 days ago #876 ↓ -6

Last evaluated 59 days ago. Current rank #876. Down 6 spots in the rankings.

Score history

84

Related listings

Snyk Agent Scan screenshot

Security / Agent Security

Open-source security scanner for AI agents, MCP servers, and agent skills. It auto-discovers installed agent components and checks them for prompt injection, tool poisoning, secrets, malware payloads, and related risks.

Xalgorix screenshot
#48 Xalgorix
90

Security / Penetration Testing

Self-hosted AI security testing platform for authorized pentesting and bug bounty workflows, with a local web UI, live agent telemetry, verified findings, and branded PDF reports.

Bright Security Agent screenshot

Security / Application Security

GitHub Marketplace app from NeuraLegion that scans apps and APIs for vulnerabilities, proposes fixes, and validates remediations inside GitHub workflows.

Belay screenshot
#214 Belay
88

Security / AI Agent Security

Belay is an open-source, local-first security layer for AI coding agents. It blocks dangerous commands, secret leaks, prompt injection, and risky MCP tool calls, with human approval flows and support for multiple agents.