Open-World Self-Evolution for LLM Agents — agents that build both their skills and their own verification signals from scratch, with no target-task supervision. (Code coming soon.)
Multi-agent research automation framework for LLM agents, with adversarial lab meetings, paper-review rounds, auditable Markdown workflows, an autonomous runtime watchdog, and a pixel-art web dashboard.
LLM agents as your hyperparameter optimizer.
The python library for research and development in NLP, multimodal LLMs, Agents, ML, Knowledge Graphs, and more.
SkillOpt is a text-space optimizer that trains reusable natural-language skills for frozen LLM agents through trajectory-driven edits, validation-gated updates, and deployable best_skill.md artifacts.
🤖 Create agentic apps in a second with your prompts. Everything you need to create an LLM Agent - tools, prompts, frameworks, and models - all in one place.
A framework for discovering, compiling, and validating reusable skills for scientific agents.
OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards

Build and run agents you can see, understand and trust.
🦖 Serverless AI Agent Framework with Geo-distributed Edge AI Infra.
The World's First Virtual Terminal for AI Agents
A Python library for building AI agents that leverage the full power of Google Antigravity.
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning

OWASP Foundation web repository
AI agents and Nix: parametrable skills/instructions and tools, packaged together in a reproducible and modular fashion
Build an agent harness and control it end-to-end. Open-source SDK for production AI agents in Python & TypeScript - any model, any cloud.
Ultra-lightweight, open-source, self-hosted personal AI agent framework in Python with WebUI, tools, memory, MCP, multi-agent workflows, automation, and chat apps
The open-source memory and observability layer for AI agents — persistent memory, loop detection, hash-chained audit trails, and a live dashboard, automatic on pip install.
Memori is agent-native memory infrastructure. A LLM-agnostic layer that turns agent execution and conversation into structured, persistent state for production systems. Built for enterprise, Memori works with the data infrastructure you already run, no rip-and-replace, and deploys across managed cloud, single-tenant cloud, VPC, and on-premises.