Sandbox
35 repos for agentic-skill · Harnesses · Any agentClear
ashutoshsinghpr7/
wikiskill

WikiSkill (arXiv:2608.27454) for Hermes Agent — self-evolving agent skills via a persistent knowledge wiki. Faithful Algorithm 1 implementation with real agent runs, isolated skill gating, and a documented live run log.

144
adewale/
skill-eval-harness

Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

73
getnao/
sylph
getnao/sylphHarnesses

The open-source company brain. Run your entire company with AI agents, skills, and a self-improving context.

196

A symbiotic AI agent that remembers everything, challenges you, and extends your cognition.

737
deepklarity/
harness-kit

A kit for building with AI agents and also the engineering patterns around it.

97
sami-mag07/
scraping-agent-skeleton

Research and scraping agent skeleton: tool loop, loadable skills, fallback chains. No data included, configured via .env.

36
AlekseiUL/
openclaw-superagent

Complete AI Agent System for OpenClaw — memory, self-healing, self-improvement, voice, automation

41

Opinionated AI coding agent and dev environment automation for macOS

129
jjmartres/
ai-coding-agents

Single source of truth for AI coding agent configuration — skills, commands and rules shared across OpenCode and Pi.

45
zby/
commonplace

The theory of LLM wikis, running as one. A framework for agent-operated knowledge: typed, linked, review-gated markdown your agents execute.

88
kevinluosl/
deepbot

DeepBot is a system-level AI assistant built for both personal productivity and enterprise workflows — one-click setup, seamless experience, and native Feishu integration.

2.3k
itseffi/
agentic-os

Agentic personal OS to automate high-leverage workflows with Codex, Claude Code, Pi, OpenClaw and other coding agents/ runtime platforms.

111
mvschwarz/
openrig

Multi-agent harness that runs Claude Code and Codex together as one system

66
AlekseiUL/
agentforge-openclaw

Create production-ready skills and agents for OpenClaw. 4-level memory, auto-improvement, 22 battle-tested pitfalls.

43
postmelee/
hyper-waterfall

A human-governed AI coding workflow that distills ephemeral session context into persistent project memory—making work traceable, reviewable, and resumable.

80
joe960913/
Jixu

Durable single-Agent Harness for TypeScript: recoverable Threads, context continuity, explicit side effects, and a native TUI.

100
shamspias/
reins

A lightweight agent harness you bolt onto your app so an LLM can operate it — safely, and cheaply.

92
skillberry-ai/
cap-evolve

Optimize any AI agent’s skills, tools/MCP, and prompts against your own evals.

56
AIScientists-Dev/
Flowtrace

Run a task with AI as a flow of steps you keep, reuse, and refine, not a one-off chat.

486
shinpr/
agentic-code

Agentic coding framework powered by AGENTS.md: systematic, test-first workflows with quality gates for Cursor, Codex, Gemini CLI, and AI coding agents.

49
FlyFission/
nuclear-grade-context-engineering

AI agents now operate with authority. Authority without discipline is how complex systems fail. Nuclear’s control loop, ported to AI-assisted software engineering.

33
ReinaMacCredy/
maestro

Local-first coordination for human and agent work: durable work, decisions, dispatches, evidence, and prompt-first methods, powered by TypeScript and Bun.

232

A structured 3-agent AI dev team — Architect, Builder, Reviewer. Built from production use. Token-optimized. Works with Claude Code, VS Code, Cursor, and any AI that supports context files.

949