Sandbox
64 repos for llm-agents · ResearchClear
LycheeMem/
LycheeMem

Lightweight Long-Term Memory for LLM Agents.

1.1k
OpenLAIR/
OpenSkill
OpenLAIR/OpenSkillFrameworks & SDKs

Open-World Self-Evolution for LLM Agents — agents that build both their skills and their own verification signals from scratch, with no target-task supervision. (Code coming soon.)

88
LiXin97/
agora-lab
LiXin97/agora-labFrameworks & SDKs

Multi-agent research automation framework for LLM agents, with adversarial lab meetings, paper-review rounds, auditable Markdown workflows, an autonomous runtime watchdog, and a pixel-art web dashboard.

49
dukesun99/
Corpus2Skill

Official Findings of EMNLP 2026 implementation of Corpus2Skill: compile a document corpus into a navigable skill hierarchy that LLM agents explore at query time, with document lookup instead of a serving-time vector-search service.

85

SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal

3.8k
lmwilki/
civ6-mcp

An MCP server that lets LLM agents play Civilization VI.

172
NPC-Worldwide/npcpyFrameworks & SDKs

The python library for research and development in NLP, multimodal LLMs, Agents, ML, Knowledge Graphs, and more.

1.5k
yzfly/
Mind-Cloning-Engineering

MCE: Clone Human Souls with LLM Native Agent Skills | 基于 LLM Agent Skills 的心智克隆工程 | Agent Skills | Mind Skills | Mind Clone

60
oneal2000/
SR-Agents

SRA-Bench and SR-Agents: a benchmark and toolkit for skill-retrieval-augmented LLM agents.

103

A self-hosted sandbox for red teams to test payloads against modern detection before deployment. MCP integration lets an LLM agent drive analysis end to end.

1.5k

MCP server for the Live Tennis API — give Claude, Cursor and other LLM agents real-time tennis scores, odds and model win-probability

158

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works with Claude Code, Codex, OpenClaw, or any LLM agent.

16k
MemTensor/
skills-vote

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

301
huggingface/
funes

Durable, searchable memory of your past agent sessions.

361
ma-compbio-lab/
SkillFoundry

A framework for discovering, compiling, and validating reusable skills for scientific agents.

39
agentscope-ai/
OpenJudge
agentscope-ai/OpenJudgeFrameworks & SDKs

OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards

826
iBlinkQ/
project-cairn

Turn project work into reusable knowledge — an AI-agent skill for Claude Code & Codex

226
yomorun/yomoFrameworks & SDKs

🦖 Serverless AI Agent Framework with Geo-distributed Edge AI Infra.

1.9k
amiable-dev/
llm-council

The LLM Council works together to answer your hardest questions

41
deltadbu/
WIT-skill

Writing is thinking! WIT provides skills for scientific thinking and writing!

35
jinzcdev/
markmap-mcp-server

An MCP server for converting Markdown to interactive mind maps with export support (PNG/JPG/SVG).

285