Sandbox
64 repos for llm-agent · ResearchClear
Gen-Verse/
Skill-Entropy-RL

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning

38

A skill library that teaches Claude Opus the working disciplines of Claude Fable 5. Move the checkpoints, not the capacity.

30

A curated list of open source GitHub repositories related to ChatGPT, the OpenAI API, and Codex. Searchable via Claude Code skills.

3.2k
JingxuanC/
causal-memory

Causal memory layer for AI agents — MCP server that records decision→outcome relationships. Survives compaction.

71
evo-hq/evoPlugins

turns your codebase into an autoresearch loop — discovers what to measure, instruments the benchmark, then runs tree search with parallel subagents.

1.4k
VILA-Lab/
FigMirror

An Automated AI Agent Tool for Plotting Your Data in Any Paper's Figure Style.

510
nablo-io/
lerim

Compiles AI agent traces and truns them into reusable context.

97
cafferychen777/
ChatSpatial

MCP server for spatial transcriptomics analysis through natural language interfaces.

44
LLMQuant/
awesome-trading-agents

Curated list of LLM-driven trading agents, MCP servers, and agent skills for market research, strategy, and execution.

448
jdrhyne/
agent-skills

A collection of AI agent skills for Clawdbot, Claude Code, Codex

241
AIScientists-Dev/
Flowtrace

Run a task with AI as a flow of steps you keep, reuse, and refine, not a one-off chat.

486

A curated list of resources for Japanese natural language processing (NLP): Python libraries, LLMs, dictionaries, corpora, and datasets. Includes Claude Code skills to search resources.

1k

The open source, no-code MCP Server for AI-Native API Access

121
Tencent/
SkillHone

Continual agent skill evolution through persistent decision history. Whole-skill optimisation (SKILL.md + scripts + references) with every decision landing as a local Git issue / PR / wiki. Runs on any agentskills.io runtime — Claude Code, Codex, OpenClaw, Hermes.

146

Research-grade investment decision engine for AI agents: isolated multi-agent committee, auditable verdicts, backtests with lookahead protection, published negative results

83
webfuse-com/
awesome-autoresearch

A curated list of autonomous improvement loops, research agents, and autoresearch-style systems inspired by Karpathy's autoresearch.

2.5k
XuanRanL/
loamwright-SEO-Skill

Production-grade SEO + GEO content factory for Claude Code — research, write, fact-check, optimize, publish to WordPress, monitor. Battle-tested by Loamwright 沃匠 SEO agency.

49
Xiangyue-Zhang/
auto-deep-researcher-24x7

🔥 An autonomous AI agent that runs your deep learning experiments 24/7 while you sleep. Zero-cost monitoring, Leader-Worker architecture, constant-size memory.

1.3k
ashutoshsinghpr7/
wikiskill

WikiSkill (arXiv:2608.27454) for Hermes Agent — self-evolving agent skills via a persistent knowledge wiki. Faithful Algorithm 1 implementation with real agent runs, isolated skill gating, and a documented live run log.

144

Distilly — Distill how they think into reusable Skills for any Agent or Bot. Formerly Colleague Skill(原同事 Skill).

25k

Build AI agents that actually do things. Synapse is an open-source platform for creating, connecting, and orchestrating AI agents powered by any LLM — local, cloud or CLIs.

320
YizhiSong/
FriesTrader

Robinhood Agentic Trading agent — a fully automated AI trading bot placing real orders through Robinhood's Agentic Trading MCP, under mechanical, auditable risk rules the model cannot override. Able to run unattended on Claude Pro, no metered API spend. Not financial advice.

157