Sandbox
963 repos for code · TestingClear

Find and repair substance defects in AI-assisted prose, code, docs, and agent output. Reports defects, never authorship. Structural tests over model judgement, because LLM judges agree with human slop labels at chance.

48

Plinth is an AI-native engineering toolkit for modern Java enterprise SDLC, built around reusable Commands, Agents, Skills, and MCP Servers.

438
Evol-ai/
SkillCompass

Evaluate agent skill quality. Find the weakest link. Fix it. Prove it worked.

216

The missing DevTools for Claude Code — inspect session logs, tool calls, token usage, subagents, and context window in a visual UI. Free, open source.

3.9k
nxtg-ai/
forge-orchestrator

Forge Orchestrator: Multi-AI task orchestration. File locking, knowledge capture, drift detection. Rust.

159
Da7-Tech/
SureForge

Agent Skill for complex work: research before asking, ask before planning, plan before building, verify before delivering, independent review before calling it done. Plain text, no runtime.

84
dcouple/
Pane

Terminal-first, open-source AI agent manager for any CLI agent (agent agnostic), any OS (mac, windows, linux). The Open-Source Agentic Development Environment for running multiple coding agents in parallel. Run locally or self-host Remote Pane to manage agents from desktop or phone. Simplify multi-agent orchestration with the runpane CLI.

454
arpitnath/
claude-capsule-kit

A toolkit that makes Claude Code better at engineering — session memory, dependency analysis, large file navigation, 18 specialist agents, and crew teams for parallel multi-branch work.

90
drvoss/
everything-copilot-cli

The definitive guide & configuration system for GitHub Copilot CLI — agents, skills, rules, multi-AI orchestration, and more

46
masteranime/
n8n-claude-skills

Production Claude Code skills for n8n from a Verified Creator's 100+ workflows

32

🥷 A code-first, strongly-typed build automation system for Deno.

38

AI coding agent with one Python core and three front-ends — headless CLI, Textual TUI, and an Electron desktop. Works with any OpenAI-compatible API, with risk-tiered permissions, event-sourced replayable sessions, and a fail-closed OS-level sandbox.

119
revfactory/
harness

A meta-skill that designs domain-specific agent teams, defines specialized agents, and generates the skills they use.

9k
jfrog/
boost

Save tokens. Maximize context, Safely

473
morluto/flameoxConnectors

Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.

121
hongnoul/hwatuConnectors

Fast, interruptible verification browser for AI coding agents: 35 ms checks, pixel diffs, live human hand-off

80

Turn tracker tickets into autonomous agent sessions

181
tikalk/
adlc-team-skills

🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

133
sbusso/
claudeclaw

Use Claude to orchestrate agents like OpenClaw

192
coleam00/
dark-factory-experiment

A repository that ships its own code. AI workflows triage issues, implement them, review, and auto-merge with no human reading the diff. Runs on Archon. The app it maintains is a cited RAG chat over YouTube transcripts, live at chat.dynamous.ai.

151
lookfree/
cc-harness

Desktop workbench for Claude Code — live subagent topology, token cost breakdown with drill-down, hook sandbox, config audit. Reads your session files locally. Electron, MIT.

48

This is MCP server for Claude that gives it terminal control, file system search and diff file editing capabilities

9.5k
open-gitagent/
opengap

A framework-agnostic, git-native standard for defining AI agents

2.9k