Vigilante is a sandbox-first orchestration layer for coding agents. It isolates every task in a git worktree, enforces strict credential scoping, and gives you full audit logs — so your agents can't burn down production.

A provider-agnostic scaffolding kit for running structured multi-agent workflows in your codebase.
Claude Code for Financial Market
A portable memory protocol for AI agents — load it as standing rules; a curation discipline + reference spec + optional cap hook.
Observability and enforcement for AI agent harnesses. Capture every run and runtime reliability with policy enforcement. 40 built-in policies, a local dashboard, no account required with a generous free cloud plan
A human-governed AI coding workflow that distills ephemeral session context into persistent project memory—making work traceable, reviewable, and resumable.
VexJoy AI Agent with Intelligent Routing - /do routes plain-English requests to the right specialist agent and gates the work with reviews, tests, and a learning loop.
Open-source self-improving QA agent for software teams. A test harness with memory. Write tests in natural language for web and mobile. agent-qa learns from every run, adapts to UI changes, and catches regressions before you ship.
Solweaver: a Codex software team with GPT-5.6 Sol orchestrating Terra and Luna workers.
A realtime voice runtime that keeps Agents talking, working, and present. Real-time Voice Runtime for AI Agents
Protocol-layer harness for DeepSeek: Python witness stack — posterior verification that keeps the protocol honest. dsh doctor --node probes included.
💾 MOI: agent-agnostic generative UI workspace. Give your agent a skill and let it build software around itself.
Nexent is a zero-code platform for auto-generating production-grade AI agents using Harness Engineering principles — unified tools, skills, memory, and orchestration with built-in constraints, feedback loops, and control planes.
A kit for building with AI agents and also the engineering patterns around it.
Harness engineering applied to knowledge production: a self-evolving multi-agent newsroom that turns your documents into a cross-linked markdown wiki. A "reground" loop pulls published pages back in before they go stale — writer ≠ reviewer, local-first, a structured alternative to RAG.
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
Durable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.
Resource-aware multi-agent orchestration for Codex and DeepSeek Harness (All in Flash DSH plugin)
Public results and task definitions for FrontierHarness Eval
Local-first, self-hosted AI agent runtime and MCP bridge with sandboxed sessions, memory, credentials, audit/replay, and a local Console.
A lightweight agent harness you bolt onto your app so an LLM can operate it — safely, and cheaply.
Let Your AI Play Detroit:Become Human
Self-hosted agent OS with skills, workflows, MCP, and second brain storage.