
Runtime intelligence system that makes MCP servers debuggable, testable, and safe to run in production.

Runtime intelligence system that makes MCP servers debuggable, testable, and safe to run in production.
Visual testing tool for MCP servers
MCP Server for interacting with Android Devices.

Model Context Protocol (MCP) Client for Apify's Actors
BrowserStack's Official MCP Server
A native desktop application for developing, testing, and debugging Model Context Protocol servers.
SmartBear's official MCP Server
Vibetest MCP - automated QA testing using Browser-Use agents
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
A test runner for agentskills.io-style AI agent skills
An MCP server that lets LLM agents play Civilization VI.
Android in docker solution with noVNC supported, video recording and mcp server
Token-efficient, local-first CLI tools for coding agents - compact Maven, npm/Node, and Go test output plus reusable development helpers.
An agent skill focused entirely on Swift Testing, helping you write better tests, migrate from XCTest, improve test architecture, and adopt modern Swift testing patterns with confidence.
Java test automation framework for web, mobile, API, CLI, database, and desktop E2E testing with a fluent API and built-in reporting.

BitDive Model Context Protocol (MCP) server. The Autonomous Quality Loop for AI agents. Provides real runtime context, before/after trace comparison, and integration testing workflows.

Self-hosted AI agent harness in a single Go binary — writes, sandbox-tests and repairs its own tools, and lets Claude Code, Codex and any MCP client build and share them.
A local-first Playwright test automation toolkit for AI coding agents (Cursor, Claude Code, Copilot), with MCP tools for codegen, browser control, and recordings.

Playwright AI Agent POM MCP ServerPlaywright AI Agent using Page Object Model (POM) architecture with MCP Server integration for automated web and mobile testing
AgentVitals Checkup (/checkup) — an AI agent skill that gives your agent a professional health checkup: dual-axis Stability + Welfare scoring, a personality-style title, and a public cross-platform leaderboard. One-line install for Claude Code / OpenClaw / Codex / Coze. Bilingual EN/ZH.

MockServer is an HTTP(S) mock server and proxy for testing that lets you mock APIs, inspect and modify live traffic, and inject failures. It supports HTTP/1.1, HTTP/2, gRPC, WebSockets, TCP and more on a single port, with additional support for HTTP/3, message brokers, and AI/LLM APIs.
Playwright for coding agents. Benchmark Claude Code, Codex, Gemini, and OpenCode on your own tasks - and test that your skills, MCP servers, and CLIs work when an agent uses them. Sandboxed YAML suites, activation checks, A/B experiments, CI gates.