Sandbox
174 repos for test · Tools · Claude CodeClear
pinecone-io/
cultivar

Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.

37
PramodDutta/
qaskills

QA Skills Directory QA Skills is a curated directory of testing-specific skills for AI coding agents (Claude Code, Cursor, Copilot, etc.).

222
KryptosAI/
mcp-observatory

CI-native security testing for MCP servers. Attack simulation, schema drift detection, and health scoring before agents depend on them.

146
heal-dev/
heal-playwright-tracer

Open-source statement-level Playwright tracer, purpose-built for AI agents. Analyzes test runs with increased accuracy.

45

Tools for AI agents to test, fix and optimise your codebase

46
stijnswapped/
Myrqen

Security testing that runs inside the coding agent you already use. Source-available, not open source.

158
MCPJam/
inspector

Testing and evaluation platform to chat, inspect, and debug MCP servers, MCP apps, and ChatGPT apps.

2.2k
callstack/
agent-device

Mobile app automation and verification for AI coding agents. CLI, MCP server, and typed Node.js API for iOS, Android, HarmonyOS, TV, web, macOS, and Linux.

4.5k
everr-labs/
everr

CLI and telemetry system for querying local, CI, and production runtime data.

45
paleo/
alignfirst

Companion products for AI-assisted software work.

86
redhat-et/
ripwire

The ripgrep of AI context: a zero-dependency C++23 CLI + MCP server for coding agents. Find what you want without reading the repo, then check you built what you meant — blast radius, tests-to-run, quality deltas. Signatures at 74.7% fewer bytes than bodies; every guess labelled, every loss published. Paddle out with a map.

1.9k
rwliebs/
Dossier

Visual planning and context control for shipping big apps with AI coding. Build from scratch or map existing repos. Dossier maps user workflows, sets agent context per feature, builds, tests and ships from one interface.

89

Bring Claude Code, Codex, and your favorite CLI agents into one visual workspace. Run agents in parallel and build executable workflows in isolated Git worktrees. Build your own AI coding team, and turn builds, tests, and dev servers into reusable canvas workflows.

46
ykdojo/
safeclaw

The easiest way to run multiple Claude Code sessions, each in its own container, with a dashboard to manage them all. Quick setup with battle-tested sensible defaults and skills.

183
EliasOenal/
term-cli

Interactive terminals for AI agents, built for what you can't --yes away. SSH+MFA, GRUB/U-Boot, debconf installers, SOL/serial consoles, fsck, cryptsetup, pdb/gdb, apt, certbot, pwsh and even Vim in tmux-backed sessions. Agent-driven, human-assisted for secrets/MFA. Single-file Python. Agent Skill. CI with 700+ tests. BSD License.

102

Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.

22k
itsmostafa/
gskill

CLI that learns repository-specific Claude skills with evolutionary search.

33
clay-good/
OpenLore

Deterministic, local-first memory and guardrails for AI coding agents with no LLM in the hot path.

303
chatml/
chatml

AI Agent Orchestrator for Claude Code

49
alibaba/
skill-up

An evaluation and evolution tool for Agent Skills.

880
antasphere/
clave

A macOS desktop app for managing multiple Claude Code sessions

47