A collection of structured AI agent skills that enable Claude Code, Cursor, GitHub Copilot, and other AI coding assistants to create, operate, debug, and govern Harness CI/CD workflows through natural language.
File-backed workflow harness for reliable Claude Code and Codex sessions.
My personal directory of skills.
Our agent harness: Skills, plugins, hooks, and utilities to improve the quality of your agent.
Agent-first local harness for OKF-compatible LLM Wikis.
Awesome list for AI agent harness engineering: tools, patterns, evals, memory, MCP, permissions, observability, and orchestration.
Skills for controlling Android, iOS and cloud phones
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

The harness layer for Claude Code — a reference implementation of harness engineering with hook-enforced dual review, state-machine gates that survive context compaction, and fail-closed safety where it counts. Quality gates that AI can't skip.
Claude Code AI Pipeline 2026 – PEV Framework Harness Plugin for Developers
Harness’ official MCP Server
📚 Two books on harness engineering — the design philosophies behind Claude Code & Codex: constraints, query loops, context governance, multi-agent verification. harness-books.agentway.dev
Turn Claude Code into its own Meta-Harness — a skill that evolves the scaffolding around a fixed model (memory, retrieval, context, prompts) via a native propose→score→Pareto loop. Native reimplementation of Meta-Harness (Lee et al. 2026).
A meta-skill that designs domain-specific agent teams, defines specialized agents, and generates the skills they use.
Repository-first deterministic migration and atomic Skill source for Claude Code and Codex.
Protocol-layer harness for DeepSeek: Python witness stack — posterior verification that keeps the protocol honest. dsh doctor --node probes included.
A kit for building with AI agents and also the engineering patterns around it.

A provider-agnostic scaffolding kit for running structured multi-agent workflows in your codebase.
Public results and task definitions for FrontierHarness Eval
🏆 Curated, ranked list of AI agent harnesses (100+) — plus an MCP server, llms.txt & JSON so agents can recommend them too. Rescored weekly.
Turn any repo into an agent-ready workspace for Claude Code, Codex, Cursor, and other coding agents.
Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.
Reverse-engineering Claude Code's 512K LOC TypeScript source: agent loop, tool system, permission model, Grove training pipeline, anti-distillation defense
The long-horizon computer-use harness. Run AI agents across desktop apps and the CLI for extended periods while preserving task state and making reliable progress on complex workflows. Features fresh-context execution, durable verified state, independent auditing, recoverable progress, and native Claude Code / Codex / OpenClaw integration.