Open-source statement-level Playwright tracer, purpose-built for AI agents. Analyzes test runs with increased accuracy.
Reusable agent skills for end-to-end software delivery, from requirements and implementation to review, testing, and release.
A Claude Code skill that burns tokens on demand. Stress test, inflate metrics, or just set money on fire.
Playwright for coding agents. Benchmark Claude Code, Codex, Gemini, and OpenCode on your own tasks - and test that your skills, MCP servers, and CLIs work when an agent uses them. Sandboxed YAML suites, activation checks, A/B experiments, CI gates.
Guard skills for coding agents, quality gates that catch AI-generated failure modes in code, tests, and docs
Production-ready PySpark ETL template for Databricks — medallion architecture, DABs, tests, DQX, CI/CD, and agentic development with Claude Code.
Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

Tools for AI agents to test, fix and optimise your codebase
Apple's official Agent Skills exported from Xcode 27 — SwiftUI, UIKit modernization, Swift Testing, C bounds-safety, and security hardening for AI coding agents.
Security testing that runs inside the coding agent you already use. Source-available, not open source.
Organisation for Claude Code inspired by time-tested Royal Navy operating procedures.

AI agents can generate code, but they still struggle to understand what they build. Reticle gives them runtime perception of web & desktop applications.
A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
Skills for AI coding agents — Laravel, PHP, React, TypeScript, testing, security, and code quality.
General-purpose Playwright automation for coding agents
Security testing toolkit for AI Agent: curated SecLists wordlists, injection payloads, and expert agents for authorized pentesting, CTFs, and bug bounties
Agent skills for Odoo addon development and OCA module migration
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
An agent skill to evolve the quality of LLM-Wiki (Graphify) at test time.
BrowserStack's Official MCP Server
Testing and evaluation platform to chat, inspect, and debug MCP servers, MCP apps, and ChatGPT apps.
The Loop Engine for Claude Code — engineer the loop, not the prompt. 1 router · 9 agents · 16 skills · 4 workflows. Fail-closed gates, test honesty, anti-anchored review.
An old coder's strategy for the agent era: don't read the code — make it run the gauntlet. Evidence-first development skill for coding agents, inspired by Uncle Bob.
Standalone engineering skills for Claude Code and Codex: review, audit, optimization, testing, product discovery, architecture, and safe publishing.