Sandbox
72 repos for evidence · Claude CodeClear
2akouwu/
reverify

Stop your AI from making things up — it proposes, deterministic tools decide, every claim checked against ground truth with evidence. Grounded facts and context survive resets. Reverse engineering is the proving ground. MCP server + CLI.

1.1k

German AI Text Humanizer for Claude Code & Codex. Audits 72 German AI-writing patterns using deterministic linters and evidence-safe rewrites. No fact-bending, no bypassing tricks.

150
DevOpsAIguru123/
awesome-agentic-devops

Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

78

Deepdive skill for Claude Code — 12-phase research pipeline: plan-review gate, parallel sub-agent search, claims-ledger triangulation with dissent protection, relevance × authority evidence filter, multi-angle red team, four-layer citation verification. 105 blocks, 29 channels, 460+ stat sources, 47 APIs, 1072 verified endpoints.

386
agent-clinic/
claude-md-doctor

Give your CLAUDE.md / AGENTS.md a checkup — audit size vitals, dead references, drifted claims, and backtest every rule against your own session history to see which rules get followed, ignored, or never used. A doctor-style report that cites its evidence.

35
haabe/
mycelium

AI made building cheap. It didn't make deciding cheap. Mycelium is a Claude Code harness that makes your agent run discovery and weigh evidence before it writes code. It earns the right to start. Built for software, courses, AI tools, and services.

45

MCP server for predictive maintenance and machinery fault diagnosis. Gives AI assistants evidence-based vibration analysis - FFT, envelope, bearing fault detection, ISO 20816-3 severity - with a measured, blind CWRU benchmark. Local-first: raw signals never leave your machine. Includes a Claude Code plugin.

83
fallow-rs/
fallow-skills

Agent skills for fallow, codebase intelligence for TypeScript and JavaScript. Teaches AI agents how to find unused code, duplication, circular deps, complexity hotspots, architecture drift, design-system drift, and (with Fallow Runtime) hot-path and cold-path evidence. Works with Claude Code, Cursor, Codex, Gemini CLI, and 30+ agents.

121
bakhtiersizhaev/
openevidence-mcp

First open-source OpenEvidence MCP server: browser-session medical research tools for Codex, Claude Code, and MCP clients

41
QinghongLin/
data2story-skill

Data Journalist Agent: Transforming Data into Verifiable Multimodal Story

155
flytohub/
flyto-core
flytohub/flyto-coreFrameworks & SDKs

AI said it finished. Flyto2 shows the proof.

480
DigitalArchivst/
Open-Genealogy

GPS-aligned AI prompts and Agent Skills for genealogical research (CC-BY-NC-SA-4.0)

73
lynxlangya/
techne

Forcing-function skills for AI agents — validated to improve behavior, not just change it.

105
cLin-c/
paper-skill

🎓 Claude Code skill — AI prompt library for academic paper writing, polishing, reviewing, translating and submitting to SCI/IEEE/Nature/TRO journals

101
boshu2/
agentops

The operations layer for agentic engineering — portable skills and contracts connecting intent, agents, software factories, and independent judgment.

434

OpenGhost is an Agent Skill for authorized web app penetration testing: Enter lab url paste credential your agent and wait everything does with help of openghost

37
pinoox/
neuromesh

The Biomimetic Context Engine & Neural Runtime for AI Coding Assistants

85
FlyFission/
nuclear-grade-context-engineering

AI agents now operate with authority. Authority without discipline is how complex systems fail. Nuclear’s control loop, ported to AI-assisted software engineering.

33
kgraph57/
paper-writer-skill

Claude Code skill for academic manuscript writing: IMRAD workflows, literature matrices, tables/figures, and tested utilities.

55

The automated approach leverages the cross-combination of high-quality papers from top conferences to uncover feasible research and innovation ideas. Through multi-level verification and convergence screening, it identifies research schemes that are feasible and have in-depth value.

89
tigerless-labs/
design-harness

Feed your agent papers and half-formed ideas — it links them into a system design you can defend. Markdown keeps the record; a visual canvas makes it readable. An Agent Skill for Claude Code & any SKILL.md-compatible agent.

217

SAW — SAFe Agentic Workflow AI Agent Harness for Multi-Agent Team Workflows Built on SAFe methodology (Scaled Agile Framework), adapted for AI agent teams (Now With AI-DLC!) Works for any team with repeatable processes: Software, Marketing, Research, Legal, Operations.

407