Use cultivar to test your Agent Skills, run them in sandboxes, and across different agents.
pentestMCP: AI-Powered Penetration Testing via MCP, an MCP designed for penetration testers.
A native desktop application for developing, testing, and debugging Model Context Protocol servers.
Touhou-inspired Agent Skills: distinct, testable, composable problem-solving workflows.
Security testing toolkit for AI Agent: curated SecLists wordlists, injection payloads, and expert agents for authorized pentesting, CTFs, and bug bounties
An agent skill to evolve the quality of LLM-Wiki (Graphify) at test time.

A self-hosted sandbox for red teams to test payloads against modern detection before deployment. MCP integration lets an LLM agent drive analysis end to end.

BitDive Model Context Protocol (MCP) server. The Autonomous Quality Loop for AI agents. Provides real runtime context, before/after trace comparison, and integration testing workflows.

Self-hosted AI agent harness in a single Go binary — writes, sandbox-tests and repairs its own tools, and lets Claude Code, Codex and any MCP client build and share them.
an open source, extensible AI agent that goes beyond code suggestions - install, execute, edit, and test with any LLM

MCP server giving AI agents full access to Julia's runtime via a live Gate — code execution, introspection, debugging, testing, and semantic search
100 field-tested Claude Code recipes for knowledge workers — prompts, steps, and 6 installable graded skills.
Product Management skills framework built on battle-tested methods for Claude Code, Cowork, Codex, and AI agents.
Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks
A Claude Code skill that encodes battle-tested editorial principles, section-specific rhetorical moves, and a structured writing pipeline for research papers. Brainstorm → Draft 0 → Evaluate → Write → Compress.
A collection of AI agent skills for Hermes Agent — research, creative, productivity, devops, and more. Each skill is self-contained and tested.
Fast, flexible, and open tooling for building intelligent workflows with Cypress.
The easiest way to run multiple Claude Code sessions, each in its own container, with a dashboard to manage them all. Quick setup with battle-tested sensible defaults and skills.
Build a Claude-Code-shaped agent harness from scratch. 7-week course, 20 chapters, ~5,000 lines of Python, 42 tests, 3 LLM providers, no frameworks.
Claude Code skill for academic manuscript writing: IMRAD workflows, literature matrices, tables/figures, and tested utilities.
Production-grade SEO + GEO content factory for Claude Code — research, write, fact-check, optimize, publish to WordPress, monitor. Battle-tested by Loamwright 沃匠 SEO agency.
NOT for educational purposes: An MCP server for professional penetration testers including STDIO/HTTP/SSE support, nmap, go/dirbuster, nikto, JtR, hashcat, wordlist building, and more.
A curated collection of top-tier penetration testing tools and productivity utilities across multiple domains. Join us to explore, contribute, and enhance your hacking toolkit!
The system of action for AI-native cybersecurity—where intent becomes governed execution, evidence becomes operational memory, and every operation improves the next.