Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.
QA Skills Directory QA Skills is a curated directory of testing-specific skills for AI coding agents (Claude Code, Cursor, Copilot, etc.).
CI-native security testing for MCP servers. Attack simulation, schema drift detection, and health scoring before agents depend on them.
Open-source statement-level Playwright tracer, purpose-built for AI agents. Analyzes test runs with increased accuracy.

Tools for AI agents to test, fix and optimise your codebase
Security testing that runs inside the coding agent you already use. Source-available, not open source.
Testing and evaluation platform to chat, inspect, and debug MCP servers, MCP apps, and ChatGPT apps.
Mobile app automation and verification for AI coding agents. CLI, MCP server, and typed Node.js API for iOS, Android, HarmonyOS, TV, web, macOS, and Linux.
Agent Skills Evaluation Framework
CLI and telemetry system for querying local, CI, and production runtime data.
Companion products for AI-assisted software work.
The ripgrep of AI context: a zero-dependency C++23 CLI + MCP server for coding agents. Find what you want without reading the repo, then check you built what you meant — blast radius, tests-to-run, quality deltas. Signatures at 74.7% fewer bytes than bodies; every guess labelled, every loss published. Paddle out with a map.
Visual planning and context control for shipping big apps with AI coding. Build from scratch or map existing repos. Dossier maps user workflows, sets agent context per feature, builds, tests and ships from one interface.
Bring Claude Code, Codex, and your favorite CLI agents into one visual workspace. Run agents in parallel and build executable workflows in isolated Git worktrees. Build your own AI coding team, and turn builds, tests, and dev servers into reusable canvas workflows.
The easiest way to run multiple Claude Code sessions, each in its own container, with a dashboard to manage them all. Quick setup with battle-tested sensible defaults and skills.
Interactive terminals for AI agents, built for what you can't --yes away. SSH+MFA, GRUB/U-Boot, debconf installers, SOL/serial consoles, fsck, cryptsetup, pdb/gdb, apt, certbot, pwsh and even Vim in tmux-backed sessions. Agent-driven, human-assisted for secrets/MFA. Single-file Python. Agent Skill. CI with 700+ tests. BSD License.
Fast, efficient, battle-tested at Alibaba's scale. Hybrid architecture code review tool: deterministic pipelines + LLM Agent, precise line-level comments, built-in multi-language ruleset (NPE, thread-safety, XSS, SQL injection), OpenAI & Anthropic compatible.
CLI that learns repository-specific Claude skills with evolutionary search.
Deterministic, local-first memory and guardrails for AI coding agents with no LLM in the hot path.
Agent Orchestration Command Center
AI Agent Orchestrator for Claude Code
A standalone CLI tool for n8n
An evaluation and evolution tool for Agent Skills.
A macOS desktop app for managing multiple Claude Code sessions