SRA-Bench and SR-Agents: a benchmark and toolkit for skill-retrieval-augmented LLM agents.
AI agent security scanner. Detect vulnerabilities in agent configurations, MCP servers, and tool permissions. Available as CLI, GitHub Action, ECC plugin, and GitHub App integration. 🛡️
Uses only the subcriptions you have. Control different agents with agents. Free, local, workspace for AI agents. Let one agent (Claude Code, Codex, Hermes, DeepSeek, Kimi, ...) spawn and drive others over MCP. No API keys, no changes to your Agent.md or Claude.md. TUI, Web UI, CLI, MCP in one package.
A test runner for agentskills.io-style AI agent skills
The AI security agent guards your code.
Real-time visualization of Claude Code agent orchestration — see your agents think, branch, and coordinate as they work.

A security scanner for your LLM agentic workflows
Computer use CLI for AI agents
Token-efficient, local-first CLI tools for coding agents - compact Maven, npm/Node, and Go test output plus reusable development helpers.
Verifiable and free cloud compute for AI agents. webMCP + MCP native. Check out our sandboxed Beta + research in the README
Run Coding Agents in Sandboxes. Control Them Over HTTP. Supports Claude Code, Codex, OpenCode, and Amp.
Terminal-first, open-source AI agent manager for any CLI agent (agent agnostic), any OS (mac, windows, linux). The Open-Source Agentic Development Environment for running multiple coding agents in parallel. Run locally or self-host Remote Pane to manage agents from desktop or phone. Simplify multi-agent orchestration with the runpane CLI.
Pluggable role definitions for AI coding agents — one command turns Claude Code / Cursor / OpenCode / Codex into a specialized professional
Open-source adversary emulation for AI agents and MCP servers.
AAS Core is the local, agent-first control plane for complete catalog discovery, agent-owned selection, stack validation, and planning, backed by 2,115+ agentic skills. Includes CLI, local MCP, catalog, plugins, and Workbench.
A terminal-first agentic development environment for agentic coding. Build for CLI/TUI agent. Runtime for Coding Agent, Tmux alternative
Agent Beacon is the world's first open-source telemetry layer for AI agents wherever they run: locally, in CI, in the browser, or in the cloud.
Mobile app automation and verification for AI coding agents. CLI, MCP server, and typed Node.js API for iOS, Android, HarmonyOS, TV, web, macOS, and Linux.
Superset is an agentic IDE to orchestrate 100+ coding agents in parallel. Run any agent with your own subscription.
TeamCity from your terminal – or your AI's. Builds, logs, agents, agent terminals, queues.
Token efficient Claude Code full Python rebuild. AI Coding Agent in 310K LoC Python.
All-in-One Sandbox for AI Agents that combines Browser, Shell, File, MCP and VSCode Server in a single Docker container.
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

Lightweight coding agent written in Rust, optimized for memory footprint and performance