Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
Personal AI agent harness powered by Claude Code. Always-on access through Discord, Telegram, or a built-in web UI.
Multi-agent orchestration for AI coding CLIs — Claude Code, Kiro, Codex, and more, coordinated in isolated tmux sessions
A standalone agent runner that executes tasks using MCP (Model Context Protocol) tools via Anthropic Claude, AWS BedRock and OpenAI APIs. It enables AI agents to run autonomously in cloud environments and interact with various systems securely.
Governance framework for AI coding agents. It runs them through a five-step workflow (plan, build, review, test, ship) where no step counts as done without evidence. Drop-in rules and guardrails for Claude Code, Codex, Cursor, Copilot, and Antigravity, via AGENTS.md.
Run multiple agents in parallel sandboxed VMs, with a single command, on your PC or in the cloud
WikiSkill (arXiv:2608.27454) for Hermes Agent — self-evolving agent skills via a persistent knowledge wiki. Faithful Algorithm 1 implementation with real agent runs, isolated skill gating, and a documented live run log.
A practical framework for AI-Assisted Research in Mathematics and Machine Learning
Open-source self-improving QA agent for software teams. A test harness with memory. Write tests in natural language for web and mobile. agent-qa learns from every run, adapts to UI changes, and catches regressions before you ship.
Portable AI agent orchestration with mechanical protocol enforcement. 186 agents, zero runtime dependencies.
ADE( Agentic Development Environment) The spec-driven environment for AI coding agents, where your planning becomes lasting, shared project context.
Durable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.
🛠️ The meta-harness for AI agents — scaffold your own focused, branded agent harness with its own npx CLI, MCP server, memory, learning loop, and witness-signed releases. Works with Claude Code, Codex, pi.dev, Hermes, OpenClaw, and RVM (hardware-isolated sandbox).

Agentic dev environment for DevPods, Codespaces & Rackspace Spot — one-command setup of Ruflo orchestration (215+ MCP tools, 60+ agents upstream).
A local multi-agent harness that works with your existing Claude Code, Codex subscriptions, allows you to run an office of agents
A lightweight agent harness you bolt onto your app so an LLM can operate it — safely, and cheaply.
a coding Agent, rpc plugin, sub-agents, hashline edits, and mcp
Vigilante is a sandbox-first orchestration layer for coding agents. It isolates every task in a git worktree, enforces strict credential scoping, and gives you full audit logs — so your agents can't burn down production.
Give AI agents ambitious work without losing the plot. nac is an open-source harness for long-running tasks, using a central orchestrator, threads, and structured episodes to stay aligned with your intent.
Multi-agent harness that runs Claude Code and Codex together as one system

Multi-Agent Harness for Production AI
a recursive self-improving harness designed to help your agents (and future iterations of those agents) succeed on any task
Turn any repo into an agent-ready workspace for Claude Code, Codex, Cursor, and other coding agents.
Open-source sandboxed agent harness for teams. Giving every employee a secured personal agent.