Sandbox
82 repos for research · Any agent · TestingClear
2389-research/
claude-plugins

28 plugins and MCP servers for Claude Code — TDD, multi-agent orchestration, iterative refinement, binary RE, structured decisions. Install any skill in one command.

95

CoCo Super Intelligence is the orchestration layer that turns Claude Code, Cursor, or Codex into an engineering department: a routed advisory board, 185 skills, 280 commands, persistent state. Local. Open-core — MIT core; Super Intelligence is proprietary, own-use.

220
lllllllama/
RigorPilot-Skills

README-first research reproduction skills with bounded execution, auditable evidence, and byte-preserving README annotations.

486
HezaoHezao/
poirot
HezaoHezao/poirotFrameworks & SDKs

Poirot is a deep research agent kernel built for those who care about how agents are architected.

217

A builder, not just a researcher. Agent skills that turn top-grossing app patterns into native-quality mobile screens.

1.3k
Tiger3807861189/
J-Space-Cognition-Suite-V3.7

J-Space Cognition Suite V3.7 - AI cognitive-enhancement Skills based on Anthropic's J-space global workspace research. | 哔哩哔哩:Tiger380 (UID 3494375382321675) — https://space.bilibili.com/3494375382321675

3k
sjkim1127/
Reversecore_MCP

A security-first MCP server that empowers AI agents to perform automated reverse engineering, malware analysis, forensics, vulnerability research, and SAST — powered by Radare2, YARA, LIEF, Capstone, and more.

201
iliaal/
whetstone

AI-powered development tools. 19 agents, 22 commands, 32 skills, 1 hook, 1 MCP server for code review, research, design, and workflow automation.

33
vje013/
darwin-agentic-cloud

Verifiable and free cloud compute for AI agents. webMCP + MCP native. Check out our sandboxed Beta + research in the README

36
appsecco/
vulnerable-mcp-servers-lab

A collection of servers which are deliberately vulnerable to learn Pentesting MCP Servers.

277

Open-source, self-hosted Claude Code - a terminal AI assistant and the Python framework behind it. Tool-calling, sandboxed execution, multi-agent teams, skills, checkpoints, unlimited context - on Pydantic AI, any model.

1.1k
bjia56/
cosmotop
bjia56/cosmotopConnectors

Multiplatform system monitoring tool using Cosmopolitan Libc

72
Epistates/
turbomcpstudio

A native desktop application for developing, testing, and debugging Model Context Protocol servers.

37
MemTensor/
skills-vote

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

301
remorses/
playwriter

Chrome extension & CLI to let agents control your browser. Runs Playwright snippets in a stateful sandbox. Available as CLI or MCP

3.9k
SawyerHood/
dev-browser

A Claude Skill to give your agent the ability to use a web browser

6.6k
flytohub/
flyto-core
flytohub/flyto-coreFrameworks & SDKs

AI said it finished. Flyto2 shows the proof.

480
adewale/
skill-eval-harness

Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

73

Touhou-inspired Agent Skills: distinct, testable, composable problem-solving workflows.

21