One-stop handbook for building, deploying, and understanding LLM agents with 60+ skeletons, tutorials, ecosystem guides, and evaluation tools.
Lightweight Long-Term Memory for LLM Agents.
Open-World Self-Evolution for LLM Agents — agents that build both their skills and their own verification signals from scratch, with no target-task supervision. (Code coming soon.)
A MCP server allowing LLM agents to easily connect and retrieve data from any database
Multi-agent research automation framework for LLM agents, with adversarial lab meetings, paper-review rounds, auditable Markdown workflows, an autonomous runtime watchdog, and a pixel-art web dashboard.
LLM agents as your hyperparameter optimizer.
SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal
An MCP server that lets LLM agents play Civilization VI.
The python library for research and development in NLP, multimodal LLMs, Agents, ML, Knowledge Graphs, and more.
SRA-Bench and SR-Agents: a benchmark and toolkit for skill-retrieval-augmented LLM agents.

A security scanner for your LLM agentic workflows
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts.

MCP server for the Live Tennis API — give Claude, Cursor and other LLM agents real-time tennis scores, odds and model win-probability
🤖 Create agentic apps in a second with your prompts. Everything you need to create an LLM Agent - tools, prompts, frameworks, and models - all in one place.

ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works with Claude Code, Codex, OpenClaw, or any LLM agent.
SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution
Durable, searchable memory of your past agent sessions.
One-stop shop for building AI-powered products and businesses with Stripe.
A framework for discovering, compiling, and validating reusable skills for scientific agents.
OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
Shell and coding agent on mcp clients

Build and run agents you can see, understand and trust.
Turn project work into reusable knowledge — an AI-agent skill for Claude Code & Codex
Practical Agent Skills — English-for-engineers coaching, Pi Agent setup, and more. Install via npx skills add.