Sandbox
212 repos for running · TestingClear

TypeScript multi-agent framework that runs in your own environment: consequential actions wait for approval and every run leaves a verifiable record. Describe the goal, not the graph. 13 built-in providers (Claude, OpenAI, Gemini, DeepSeek and more) plus any OpenAI-compatible endpoint, local models included.

6.9k

🔂 Ralph loop with PRs: Run Claude Code in a continuous loop, autonomously creating PRs, waiting for checks, and merging

1.4k

A Model Context Protocol (MCP) server that enables LLMs to run ANY code safely in isolated Docker containers.

121
risingwavelabs/
box0

Open-Source Platform for Subagents and Agent Teams. Long-running, collaborative, proactive.

81
SponsioLabs/SponsioFrameworks & SDKs

Deterministic safety solutions for probabilistic AI agents

440
Hanyuyuan6/
remote-gpu-trainer

An Agent Skill for the DL experiment lifecycle: RUN (a GPU you own or rent) → VERIFY the number is real → DELIVER reproducible, single-source figures and tables.

63
uvwt/
agentdock
uvwt/agentdockConnectors

Secure MCP runtime for AI agents to operate local machines, servers, and containers with multi-device orchestration.

784
pinecone-io/
cultivar

Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.

37
morluto/flameoxConnectors

Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.

121

The multi-agent harness that checks the work: verifies agent runs by artifacts (stop-hook gates, independent judges, append-only event logs) across Claude Code, Codex, Cursor, and 10+ runtimes.

1.3k

Local-first, self-hosted AI agent runtime and MCP bridge with sandboxed sessions, memory, credentials, audit/replay, and a local Console.

642
johannesjo/
parallel-code

Run Claude Code, Codex, and Gemini side by side — each in its own git worktree

1k
mvschwarz/
openrig

Multi-agent harness that runs Claude Code and Codex together as one system

66
evo-hq/evoPlugins

turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel subagents.

1.4k
engasnm111/
lnwjud

lnwjud — local AI-agent runtime & MCP gateway

152
avivsinai/
jenkins-cli

GitHub-style CLI for Jenkins — manage contexts, runs, logs, and admin tasks from your terminal.

87
madarco/
agentbox

Run multiple agents in parallel sandboxed VMs, with a single command, on your PC or in the cloud

390
michaelshimeles/
ralphy

My Ralph Wiggum setup, an autonomous bash script that runs Claude Code, Codex, OpenCode, Cursor agent, Qwen & Droid in a loop until your PRD is complete.

3k
termio-sh/
termio

A terminal-first agentic development environment for agentic coding. Build for CLI/TUI agent. Runtime for Coding Agent, Tmux alternative

434

An open-source plugin that runs inside Codex and lets you use Claude Code and Claude models for review, rescue, and tracked background workflows.

198

Local-first runtime for project-scoped AI coding-agent sessions, with durable state, authority boundaries, and multi-harness interoperability.

47

Emdash is the Open-Source Agentic Development Environment (🧡 YC W26). Run multiple coding agents in parallel. Use any provider.

5.7k
matlab/
matlab-mcp-server

Run MATLAB® using AI applications with the official MATLAB MCP Server from MathWorks®. This MCP server for MATLAB supports a wide range of coding agents like Claude Code® and Visual Studio® Code.

1.5k