Sandbox
266 repos for llms · TestingClear
GammaLabTechnologies/
harmonist

Portable AI agent orchestration with mechanical protocol enforcement. 186 agents, zero runtime dependencies.

2.3k

Provider-neutral control plane for durable-state agent swarms: subprocess workers, leases, artifacts, memory, and deterministic stitching.

414
vdaubry/
bottega

Coding agent orchestration for engineering teams — shipped as a spec plus a working reference implementation.

84
agentforce314/
clawcodex

Token efficient Claude Code full Python rebuild. AI Coding Agent in 310K LoC Python.

897
google/adk-goFrameworks & SDKs

An open-source, code-first Go toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.

8.8k
haddock-development/
claude-reflect-system

Continual Learning & Self-improving skills system for Claude Code - learn from corrections, never repeat mistakes

376
Sumanth077/
Hands-On-AI-Engineering

A curated collection of practical AI projects implementing OCR systems, RAG, AI agents, and other AI use cases.

3.4k
anhvt52/
jetpack-compose-skills

Agent skill for modern Android development with Jetpack Compose — best practices for code generation and review

96
uber/
ADR

ADR secures enterprise AI agents through observability, security benchmarking, and threat detection. Deployed at Uber.

1.5k
MrZoyo/
deslop-GPT

Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

125
seuros/
action_mcp
seuros/action_mcpFrameworks & SDKs

Rails Engine with MCP compliant Spec.

120
affaan-m/
ECC
affaan-m/ECCHarnesses

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

258k
JuliusBrussee/
caveman

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

105k
1 add
evo-hq/evoPlugins

turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel subagents.

1.4k
yohey-w/
multi-agent-shogun

Samurai-inspired multi-agent system for Claude Code. Orchestrate parallel AI tasks via tmux with shogun → karo → ashigaru hierarchy.

1.4k
joe960913/
Jixu

Durable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.

100

Tools for AI agents to test, fix and optimise your codebase

46
smithersai/smithersFrameworks & SDKs

Smithers is an agentic workflow framework for defining workflows in simple TypeScript configuration files and executing them quickly, durably, and reliably

410
HenryZ838978/
deepseek-harness

Protocol-layer harness for DeepSeek: Python witness stack — posterior verification that keeps the protocol honest. dsh doctor --node probes included.

49

Universal mobile devtool for Agents & Humans - control iOS Simulators, Android Emulators, and real devices from a single dashboard and CLI

368
Sahir619/
fable-method

The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

2.3k