Sandbox
@marmotdata/marmot

MCP server for data catalog metadata

Marmot is a data catalog that turns tables, topics, queues, APIs, and dashboards into context that agents can read. It exposes that context through MCP, plus an API and UI, so agents can work from real metadata instead of partial guesses.

613 stars27 forksGoUpdated 7d ago
Marmot - The Open-Source Data Catalog for Agents and Engineers
Marmot Data728 views • 5 months ago
Who it's for

Builders who want their agents to inspect data assets, lineage, and ownership before they act.

What it delivers

You can give your agent certified metadata about your data stack instead of re-explaining the same context in every session.

What it does

Search across assets

Find data assets with full-text search, structured queries, boolean logic, and metadata filters.

Interactive lineage

Trace data flow from source to destination and check impact before making changes.

Metadata for many asset types

Store context for tables, topics, queues, APIs, and dashboards in one place.

Team collaboration

Assign ownership, add business context, and maintain shared glossaries.

Agent access through MCP

Expose catalog context to AI agents through MCP, the API, and the UI.

Plugin-based sources

Extend the catalog with plugins for systems like BigQuery, Kafka, Postgres, OpenAPI, and many more.

README

Marmot

Discover any data asset in seconds. Then let your AI do the same.

The open-source context layer for your AI. Catalog your tables, topics, queues and APIs then expose real metadata to your AI agents.

DocumentationDeployCommunity

What is Marmot?

Marmot is an open-source data catalog for teams who want powerful data discovery without enterprise complexity. Catalog every data asset, enrich it with the context that matters and make it accessible to your team and your AI tools.

Unlike traditional catalogs that require extensive infrastructure and configuration, Marmot ships as a single binary with an intuitive UI, making it easy to deploy and start cataloging in minutes.

Features

  • Search everything: Find any data asset in seconds with full-text search plus structured queries, boolean logic and metadata filters.
  • Interactive lineage: Trace data flows from source to destination and analyse impact before making changes.
  • Metadata-first: Store rich metadata for any asset type, from tables and topics to APIs and dashboards.
  • Team collaboration: Assign ownership, document business context and maintain shared glossaries.
  • AI-ready: Expose certified context through MCP, the API and the UI.
Marmot context layer showing plugins, integrations and AI agents

Deploy

New to Marmot? Follow the Deploy documentation for a guided setup.

Development

See Local Development for how to get started developing locally.

Community

Join our Discord community for help, feedback and updates on new features.

Contributing

All types of contributions are encouraged and valued!

  • Report bugs or suggest features via GitHub Issues
  • Improve documentation
  • Build new plugins for data sources

Before contributing, please check out the Contributing Guide.

FAQ

What is Marmot?

Marmot is an open-source data catalog for teams who want powerful data discovery without enterprise complexity. It catalogs every data asset (tables, topics, queues, APIs), enriches it with context, and makes it accessible to both your team and AI tools. Unlike traditional catalogs requiring extensive infrastructure, Marmot ships as a single binary with an intuitive UI.

Why Use Marmot?

FeatureBenefit
Search EverythingFind any data asset in seconds with full-text search, structured queries, boolean logic, and metadata filters
Interactive LineageTrace data flows from source to destination, analyze impact before making changes
Metadata-FirstStore rich metadata for any asset type (tables, topics, APIs, dashboards)
Team CollaborationAssign ownership, document business context, maintain shared glossaries
AI-ReadyExpose certified context through MCP, API, and UI
Single BinaryNo complex infrastructure required, deploy in minutes

How Does Marmot Integrate with AI Agents?

Marmot exposes certified context through MCP (Model Context Protocol), enabling AI agents to query real metadata about your data assets. This allows AI tools to understand your data landscape, trace lineage, and make informed decisions based on actual metadata rather than guessing.

What Data Sources Can I Catalog?

Marmot supports cataloging various data asset types through its plugin system:

  • Tables (databases, data warehouses)
  • Topics (message queues, event streams)
  • Queues (job queues, messaging systems)
  • APIs (REST, GraphQL, internal services)
  • Dashboards (visualization tools, BI platforms)

How Do I Deploy Marmot?

Quick Start Options:

MethodDescription
Documentation GuideFollow the Deploy guide for step-by-step setup
Single BinaryDownload and run the single binary for your platform

Is Marmot Free?

Yes! Marmot is open-source software licensed under the MIT License. You can use, modify, and distribute it freely. Self-hosting is completely free with no licensing costs.

How Can I Contribute?

All contributions are welcome:

  • Report bugs or suggest features via GitHub Issues
  • Improve documentation (README, guides, API docs)
  • Build new plugins for additional data sources
  • Check the Contributing Guide before contributing

Where Can I Get Help?

ResourceLink
Documentationmarmotdata.io/docs
Discord CommunityJoin Discord
GitHub IssuesReport issues

License

Marmot is open-source software licensed under the MIT License.

Files in the repo

Repository payload28 top-level entries
  • .github
  • charts
  • cmd
  • docs
  • internal
  • pkg
  • plugins
  • sdk
  • web
  • .dockerignore
  • .gitignore
  • .goreleaser.yaml
  • .prowlabels.yaml
  • CONTRIBUTING.md
  • Dockerfile
  • Dockerfile.goreleaser
  • Dockerfile.goreleaser.distroless
  • go.mod
  • go.sum
  • install.sh
  • LICENSE
  • Makefile
  • marmot.svg
  • OWNERS
  • README.md
  • server.json
  • SKILL.md
  • Tiltfile

Discussion (0)

Ask about usage, or say what you built with it

Sign in to join the discussion.

No comments yet. Be the first to say what this is good for.

More connectors

Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface

86k

High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph — average repo in milliseconds. 158 languages, sub-ms queries, 99% fewer tokens. Single static binary, zero dependencies.

43k

Universal provider proxy for OpenAI Codex & Claude Code — use any LLM (Claude, Gemini, Grok, DeepSeek, Ollama…) with Codex CLI, App, SDK, and Claude Code

14k
okf-memory/
okf-agent-memory

Git-native persistent memory for AI coding agents. Implements Google OKF v0.2 with sub-300µs in-memory BM25 search, embedded MCP server, and progressive disclosure. Slashes token bloat by 80% with zero external databases or dependencies. Built in pure Go.

547
2akouwu/
reverify

Stop your AI from making things up — it proposes, deterministic tools decide, every claim checked against ground truth with evidence. Grounded facts and context survive resets. Reverse engineering is the proving ground. MCP server + CLI.

1.1k
t8y2/dbxConnectors

20 MB lightweight cross-platform database client for 90+ databases, including MySQL, PostgreSQL, SQLite, Redis, MongoDB, DuckDB, SQL Server, and Dameng. Built-in AI, MCP Server, CLI, desktop and Docker. | 轻量级跨平台数据库管理工具,支持 MySQL、PostgreSQL、SQLite、Redis、MongoDB、达梦等 90+ 数据库,提供桌面端、Docker、CLI、内置 AI 助手和 MCP Server。

19k