A Model Context Protocol (MCP) server that enables LLMs to run ANY code safely in isolated Docker containers.
Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.
AI coding agent with one Python core and three front-ends — headless CLI, Textual TUI, and an Electron desktop. Works with any OpenAI-compatible API, with risk-tiered permissions, event-sourced replayable sessions, and a fail-closed OS-level sandbox.
Agent workflow for deadline-bound execution with evidence gates
Durable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.
Standalone MCP server for cross-pane AI agent communication via tmux. Lets Claude Code, Gemini CLI, Codex, and Kimi CLI talk to each other through tmux panes.
Non-Invasive goroutine inspector
Orchestrate AI coding agents (Claude Code, Codex) as parallel subagents over tmux — a loop-engineering runtime with auto-continue, execute-then-review, and cross-session memory.
Code graphs, wikis, and research-backed agentic coding. Deterministic tools for nondeterministic workflows, in one opinionated Claude Code config.
Composable agent runtime with enforced isolation boundaries
The first AI plugin that speaks first. Code-enforced learning + active forgetting + PAC (Proactive Accountability Challenge). Works with Claude Code, Gemini CLI, Hermes, OpenClaw.

The environment engine for agent worktrees - zero-config dep sync + port/DB/compose isolation
MCP server for guiding Coding Agents via end-to-end requirements to implementation plan pipeline
[COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.
Open-Source Platform for Subagents and Agent Teams. Long-running, collaborative, proactive.
Agent Skills marketplace: framework-aware skills for code review, documentation, test-plan generation, AI-writing detection, architectural analysis, and git workflows — for Python, Go, Rust, Elixir, React, Remix, and iOS/Swift. Works with Claude Code, Codex, and any agent that supports Agent Skills.
A modular external attack surface mapping tool integrating tools for automated reconnaissance and bug bounty workflows.

Atari Lynx emulator, debugger, and embedded MCP server for macOS, Windows, Linux, BSD and RetroArch.
Unified execution environment for Python code, shell commands, and programmatic MCP tool calls.
Multi Debugger MCP server that enables LLMs to interact with GDB and LLDB for binary debugging and analysis.
A complete catalog of Agent Skills (agentskills.io) for Zephyr RTOS development.
Resource-aware multi-agent orchestration for Codex and DeepSeek Harness (All in Flash DSH plugin)
Computer use CLI for AI agents
Real-time .NET proxy and dashboard for inspecting AI coding agent API calls (currently supports Claude Code)