Sandbox
1,015 repos for β€œtesting”Clear
Dave-London/
Pare

Dev tools, optimized for agents. Structured, token-efficient MCP servers for git, test runners, npm, Docker, and more.

138
FrankS-IntelLab/
agentic-kaggle-skill

πŸ€– AI Agent-driven Kaggle competition workflow. Battle-tested patterns for score stabilization, submission troubleshooting, kernel workflows, and spec-driven development.

183

MCP server for full Godot 4.x engine control: 157 tools for AI-driven game development (GDScript and C#/.NET). Tested with Godot 4.7.

454
Xopoko/
build-swift-apps

Build, debug, profile, test, refactor, and release Swift apps across iOS, macOS, Xcode, SwiftUI, SwiftPM, Tuist, and App Store Connect.

45
azalio/
map-framework

Plan-then-build AI coding for Claude Code & Codex CLI β€” you approve the plan before the model writes a line of code. SPEC β†’ PLAN β†’ TEST β†’ CODE β†’ REVIEW β†’ LEARN

156
tanbro/
uiautomator2-mcp-server

A MCP (Model Context Protocol) server that provides tools for controlling and interacting with Android devices using uiautomator2.

41
remiphilippe/
mcp-unreal

MCP server that gives AI coding agents (Claude Code, Cursor, etc.) full control over Unreal Engine 5.7 projects β€” headless builds & tests, Blueprint editing, actor manipulation, procedural mesh generation, and UE API documentation lookup.

69
7gugu/
whistle-mcp

A Whistle proxy management tool based on Model Context Protocol that allows AI assistants to directly control local Whistle proxy servers, simplifying network debugging, API testing, and proxy rule configuration through natural language interaction.

44

Battle-tested agent skills mapping the footguns of Kotlin, Compose Multiplatform and the desktop JVM β€” mined from a production music app, not from documentation.

217
Clear-Sights/
Makoto

When Claude says β€œtests pass”, makoto checks the record β€” every claim held against the agent's own logged deeds; fakes blocked, not warned. Each check ships at measured zero false positives. θͺ 

45
Zulut30/
Wordpress-skills

Professional Agent Skill for building, auditing, testing, and releasing modern WordPress plugins with Codex, Cursor, and Claude Code.

33
callstack/
agent-device

Mobile app automation and verification for AI coding agents. CLI, MCP server, and typed Node.js API for iOS, Android, HarmonyOS, TV, web, macOS, and Linux.

4.5k
shinpr/
agentic-code

Agentic coding framework powered by AGENTS.md: systematic, test-first workflows with quality gates for Cursor, Codex, Gemini CLI, and AI coding agents.

49

humanizer, but for code β€” an agent skill that removes AI-generated code slop: duplicated helpers, try-import fallbacks, broad excepts, speculative abstractions. Test-gated, behavior-preserving.

48
arpitg1304/
robotics-agent-skills

Agent skills that make AI coding assistants write production-grade robotics software. ROS1, ROS2, design patterns, SOLID principles, and testing β€” for Claude Code, Cursor, Copilot, and any SKILL.md-compatible agent.

356

BitDive Model Context Protocol (MCP) server. The Autonomous Quality Loop for AI agents. Provides real runtime context, before/after trace comparison, and integration testing workflows.

75

Self-hosted AI agent harness in a single Go binary β€” writes, sandbox-tests and repairs its own tools, and lets Claude Code, Codex and any MCP client build and share them.

508
firish/
claude_code_vs

Bring Claude Code to Visual Studio 2026: A native diff with accept/reject, a live debugger Claude can drive autonomously, Roslyn code navigation, and Test Explorer integration. The IDE half of Claude Code's integration protocol. Community-built, unofficial.

86
BaseInfinity/
claude-sdlc-harness

Self-evolving SDLC enforcement for AI coding agents β€” hooks, skills, and one-command setup for Claude Code. Plan before coding, test before shipping, escalate when uncertain. Measures itself getting better over time.

49
aaif-goose/
goose

an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM

54k
TheGreenCedar/
codex-autoresearch

A codex plugin for running optimization loops inside a codebase. It is useful when you have a measurable target and many possible changes to try: test runtime, build speed, bundle size, model loss, Lighthouse scores, memory use, query latency, or any other metric you can print from a script.

837
ianho7/
ai-friendly-web-design-skill

A skill for coding agents that build, review, and refactor Web UI that should be easier for humans, screen readers, browser automation, Playwright tests, and AI agents to understand and operate.

76