Sandbox
72 repos for evidence · Claude CodeClear

Evidence-first reading for AI agents — turn articles, books and PDFs into traceable claims, evidence, source locations and knowledge maps.

49
yha9806/
academic-writing-toolkit

Local-first, evidence-controlled academic writing workflows for AI agents, with bounded revision, clean-room review, and release governance.

38
lllllllama/
RigorPilot-Skills

README-first research reproduction skills with bounded execution, auditable evidence, and byte-preserving README annotations.

486

Fable-style spec + evidence gate for Claude Code + Codex. Makes Opus/Codex work under Fable-like discipline: blocks every edit until a deterministic spec passes, and there is no "done" without live acceptance evidence. Spec-first, verification-gated, forbidden-paths enforced.

46
GanyuanRan/
Aegis

Make AI coding agents architecture-aware: baseline-first, evidence-verified, drift-checked, and safe across long tasks.

1.2k
gotalab/
goal-setter-skill

Shape rough requests into evidence-backed /goal completion contracts — an Agent Skill for Claude Code and Codex

103

Evidence-grounded repository audit CLI - deterministic scanner, MCP server, live dashboard, and a GitHub Action that posts PR diffs.

156
yaojingang/
GEOHub

GEOHub: open, evidence-bounded GEO and SEO agent skills for AI Search, with research-grounded discovery, diagnosis, content, measurement, and one-line SEO planning.

156
tigerless-labs/
seo-ops

SEO foundation checks as an Agent Skill: give it a URL, get a crawler's-eye pass/fail report with evidence. 26 structural checks, zero LLM, deterministic.

136
millwright-labs/
minto-pyramid-skill

Agent Skill: make Claude write in Barbara Minto's Pyramid Principle - answer first, grouped reasons, evidence under each.

70
akseolabs-seo/
seo-coach

An open-source AI SEO coach for beginners: practice on your own website, make evidence-based decisions, and track verifiable progress without expensive tools.

91
Zhen-Bo/
smell-check

Agent Skill for code and test smell audits. Evidence-ranked findings from Refactoring, Clean Code, and the test-smell literature. Formerly pragmatic-code-review.

237

Tamper-evident integrity monitor for the MCP config & server files your local AI agents load.

46

Remote approvals, policy checks, and execution evidence for unattended AI agents.

301
xf686/
Meet-Reviewer-2

Meet Reviewer 2 before they meet you ! A Claude Code skill that red-teams your paper draft — a simulated peer-review panel + an evidence-grounded fix list, before you submit.

45
AaravKashyap12/
advise-project-approach

A portable project-planning skill for Codex, Claude Code, pi, Hermes, and Agent Skills-compatible harnesses. Evidence before build advice.

300
DY-2026/
GameDesignOS

Local-first game design OS for AI agents: turn sessions into evidence, experiments, reviewable decisions, and durable project memory—Human Gates and rollback.

385

Cut context bloat in your AI-agent stack: find and safely prune unused skills, MCP servers and subagents from real transcript evidence

56
PinkR1ver/
vibe-roast

Local-first AI coding personality profiler, usage dashboard, and evidence-grounded roast.

49
amplifthq/
opentag

Mention any ACP coding agent from Slack, GitHub, GitLab, Linear, or Lark. OpenTag runs Claude Code, Codex, Cursor and more on your own machine, then replies in-thread with verified, evidence-backed results.

1.4k
morluto/flameoxConnectors

Runtime evidence that helps agents trace, profile, and burn down hotspots in application and native code, GPU kernels, and inference stacks.

121
AmazingAng/
old-coder

An old coder's strategy for the agent era: don't read the code — make it run the gauntlet. Evidence-first development skill for coding agents, inspired by Uncle Bob.

720
GRCEngClub/
claude-grc-engineering

Open-source GRC toolkit from the GRC Engineering Club. Claude Code plugins for evidence collection, SCF crosswalks, multi-framework gap reports, OSCAL workflows.

398