Sandbox
918 repos for r · TestingClear
Gentleman-Programming/
gentle-pi

Turn Pi into el Gentleman: a senior-architect development harness with SDD/OpenSpec, subagents, strict TDD evidence, review guardrails, and skill discovery.

715
shinpr/
codex-workflows

Development workflows for Codex that keep technical rigor and edge-case handling from becoming product overengineering.

37
morluto/flameoxConnectors

Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.

121
ferrislucas/
iterm-mcp

A Model Context Protocol server that executes commands in the current iTerm session - useful for REPL and CLI assistance

567
Adancurusul/
serial-mcp-server

Rust MCP server and CLI for serial/UART devices, with JSON macro automation and agent skills for repeatable timed workflows.

91

Find and repair substance defects in AI-assisted prose, code, docs, and agent output. Reports defects, never authorship. Structural tests over model judgement, because LLM judges agree with human slop labels at chance.

48

Save 40%+ on agent token costs with code graphs: call graphs, dependency graphs, dead code detection, and blast radius analysis.

92
w95/
awesome-claude-corporate-skills

166 production-ready Claude AI skills organized by corporate role — executive leadership, finance, HR, marketing, sales, legal, operations, engineering, product, data, customer success, procurement & document processing

195
pacifio/
cersei
pacifio/cerseiFrameworks & SDKs

The Rust SDK for building coding agents. Tools, streaming, graph, sub-agent orchestration, MCP — as composable functions

453
TheGreenCedar/
codex-autoresearch

A codex plugin for running optimization loops inside a codebase. It is useful when you have a measurable target and many possible changes to try: test runtime, build speed, bundle size, model loss, Lighthouse scores, memory use, query latency, or any other metric you can print from a script.

837

humanizer, but for code — an agent skill that removes AI-generated code slop: duplicated helpers, try-import fallbacks, broad excepts, speculative abstractions. Test-gated, behavior-preserving.

48
MorDavid/
ExternalAttacker-MCP

A modular external attack surface mapping tool integrating tools for automated reconnaissance and bug bounty workflows.

78
addyosmani/
factory

A reference software factory for Claude Code and Codex

179
AMAP-ML/
LongHorizon-Harness

The long-horizon computer-use harness. Run AI agents across desktop apps and the CLI for extended periods while preserving task state and making reliable progress on complex workflows. Features fresh-context execution, durable verified state, independent auditing, recoverable progress, and native Claude Code / Codex / OpenClaw integration.

1.5k
artemiimillier/
bulletproof

Turns AI agents from chaotic code generators into disciplined engineers. 12-stage workflow from research to production.

151
FROWNINGdev/
django-orm-lens

Free, MIT alternative to paid Django schema review. Blast radius on every PR, schema drift, N+1 across functions, ER diagrams, MCP server. No DB, no Django boot, no Pro tier.

73
affaan-m/
claude-swarm

Multi-agent orchestration for Claude Code — decompose tasks, coordinate agents, visualize everything in a rich terminal UI

357
nxtg-ai/
forge-orchestrator

Forge Orchestrator: Multi-AI task orchestration. File locking, knowledge capture, drift detection. Rust.

159
aAAaqwq/
AGI-Super-Team

An installable, cross-framework AI organization: C-suite agents, expert subagents, curated skills, independent review, and one-command setup across 18 AI client/runtime adapters.

91
trailofbits/
skills

Trail of Bits Claude Code skills for security research, vulnerability detection, and audit workflows

7k
0xranx/
agentbrief

Pluggable role definitions for AI coding agents — one command turns Claude Code / Cursor / OpenCode / Codex into a specialized professional

45

Vibe Check is a tool that provides mentor-like feedback to AI Agents, preventing tunnel-vision, over-engineering and reasoning lock-in for complex and long-horizon agent workflows. KISS your over-eager AI Agents goodbye! Effective for: Coding, Ambiguous Tasks, High-Risk tasks

502
pbshgthm/
arc-skill

An agent skill that plays ARC-AGI-3. One rule: say what an action will do before you spend it. Claude Code on Opus 5 finished all 25 public games at 100.00 RHAE in 7,645 actions.

89
fcavalcantirj/
claude-code-eyes

Give Claude Code eyes 👁️ — a camera skill so it can SEE real hardware, displays and wiring: verify a rendered panel, check wiring before power-on, and catch bugs that live on the glass, not the logs.

44