Sandbox
32 repos for lear · Any agent · ResearchClear
cxcscmu/
SkillLearnBench

[COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.

83
BennettSchwartz/
membrane
BennettSchwartz/membraneFrameworks & SDKs

A selective learning and memory substrate for agentic systems — typed, revisable, decayable memory with competence learning and trust-aware retrieval.

95
ZIB-IOL/
The-Agentic-Researcher

A practical framework for AI-Assisted Research in Mathematics and Machine Learning

95
keyuchen21/
agentic-engineering-handbook

The definitive OpenAI, Claude, MCP, Harness, Evals, and Production Agent Systems learning roadmap.

187
appsecco/
vulnerable-mcp-servers-lab

A collection of servers which are deliberately vulnerable to learn Pentesting MCP Servers.

277
eli-labz/
Cognitive-Core-Skills

A universal, industry-neutral taxonomy of cognitive core skills (perception, memory, reasoning, planning, action, verification, learning, governance) for LLMs, SLMs, AI agents, and world models — with schemas, 159 skill cards, benchmarks, and CI.

165
Erlemar/
dswok

A collection of notes on Data Science

33
NeoLi00/
memX

memX: self-learning, self-maintaining memory plugin for AI agents; native support for claude code, codex, and openclaw

407
KatherLab/
STAMP
KatherLab/STAMPFrameworks & SDKs

Solid Tumor Associative Modeling in Pathology

128
Gen-Verse/
Skill-Entropy-RL

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning

38
wanshuiyin/
ARIS-in-AI-Offer
wanshuiyin/ARIS-in-AI-OfferTutorials & Guides

Bilingual (中文+EN) ML / LLM / diffusion / agent interview cheat sheets for AI 秋招 — generated by ARIS /interview-cheatsheet, rendered by /render-html into single-file HTML, reads anywhere — plus a CV→DBLP-fact-checked academic homepage generator and hand-authored long-form blogs 🌱

473
TokenRhythm/
opensquilla

OpenSquilla — Token-Efficient AI Agent with same budget, higher intelligence density

7k
lllllllama/
RigorPilot-Skills

README-first research reproduction skills with bounded execution, auditable evidence, and byte-preserving README annotations.

486
oneal2000/
SR-Agents

SRA-Bench and SR-Agents: a benchmark and toolkit for skill-retrieval-augmented LLM agents.

103
WecoAI/
awesome-autoresearch

Curated list of AutoResearch use cases with optimization traces and open source implementations

1k
kyegomez/swarmsFrameworks & SDKs

The Enterprise-Grade Multi-Agent Orchestration Framework. Website: https://swarms.ai

7.2k
cafferychen777/
ChatSpatial

MCP server for spatial transcriptomics analysis through natural language interfaces.

44
scitex-ai/
scitex-python
scitex-ai/scitex-pythonFrameworks & SDKs

Umbrella package for SciTeX — reproducible science from raw data to manuscript

85
Ladbaby/PyOmniTSFrameworks & SDKs

🔬 A Researcher&Agent-Friendly Framework for Time Series Analysis. Train Any Model on Any Dataset!

98