Sandbox
7 repos for pag · DataClear
sderosiaux/
chrome-agent

Web tasks that compile. Other browser tools tell your AI agent the click worked. chrome-agent reads the page back after every action and answers in JSON what actually happened. One 3 MB Rust binary, no Node, no cloud.

88
NameetP/
pdfmux

PDF extraction that audits its own output — and certifies any other extractor's, catching pages they silently dropped. Verify signed manifests offline: free, MIT, no account. 0.903 on opendataloader-bench, #2 of 8 engines. 7-tool MCP server.

82
adityaarsharma/
librecrawl-technical-seo-audit-mcp

The AI-native technical SEO crawler. Open-source MCP server for Claude / Cursor / Codex — 37 tools, 50+ checks, unlimited pages, WAF detection, ephemeral by design. Built on LibreCrawl. MIT.

39
sami-mag07/
scraping-agent-skeleton

Research and scraping agent skeleton: tool loop, loadable skills, fallback chains. No data included, configured via .env.

36
konippi/
servo-fetch

A self-contained browser engine that fetches, renders, and extracts web content as Markdown, JSON, or screenshots — no Chromium, no API key, no setup.

144
christinminor459/
OnionClaw

Provide AI agents with full Tor network access and dark web data through a zero-config OpenClaw skill or standalone tool.

237

Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data — all from Rust. CLI, REST API, and MCP server.

2.3k