Sandbox
47 repos for β€œscraping” Β· CodingClear
sami-mag07/
scraping-agent-skeleton

Research and scraping agent skeleton: tool loop, loadable skills, fallback chains. No data included, configured via .env.

36
yfe404/
web-scraper

Intelligent web scraping Claude Code skill with automatic strategy selection and TypeScript-first Apify Actor development

90
D4Vinci/ScraplingFrameworks & SDKs

πŸ•·οΈ An adaptive Web Scraping framework that handles everything from a single request to a full-scale crawl!

80k
us/crwConnectors

Fast, lightweight Firecrawl/Tavily alternative in Rust. Web scraper, crawler & search API with MCP server for AI agents. Drop-in Firecrawl-compatible API (/scrape, /crawl, /search). 2.3x faster than Tavily, 1.5x faster than Firecrawl in 1K-URL benchmarks. 6 MB RAM, single binary. Self-host or use managed cloud.

970

⚑ The Complete X/Twitter Automation Toolkit β€” Scrapers, MCP server for AI agents (Claude/GPT), CLI, browser scripts. No API fees. Open source. Unfollow people who don't follow back. Monitor real-time analytics. Auto follow, like, comment, scrape, without API. Follow Bot. Like bot. Grow your account automatically.

518

Adaptive Python web scraping toolkit + MCP server for AI agents. Self-healing selectors that survive site changes, TLS-fingerprint stealth to bypass anti-bot filters, CSS/XPath parsing, and 24 built-in scrapers, clean, structured, LLM-ready data from any URL.

200

Fast, local-first web content extraction for LLMs. Scrape, crawl, extract structured data β€” all from Rust. CLI, REST API, and MCP server.

2.3k

πŸ”₯ Official Firecrawl MCP Server - Adds powerful web scraping and search to Cursor, Claude and any other LLM clients.

7.4k
just-every/
mcp-read-website-fast

Quickly reads webpages and converts to markdown for fast, token efficient web scraping

161

Model Context Protocol Server for Mobile Automation and Scraping (iOS, Android, Emulators, Simulators and Real Devices)

6.6k
yokingma/
one-search-mcp

πŸš€ OneSearch MCP Server: Web Search & Scraper & Extract, Support agent-browser, SearXNG, Tavily, DuckDuckGo, Bing, etc.

140
antibrow/
anti-detect-browser-skills

Launch and manage anti-detect browsers with unique real-device fingerprints for multi-account operations, web scraping, ad verification, and AI agent automation.

218

A powerful Model Context Protocol (MCP) server that provides an all-in-one solution for public web access.

2.6k
web-agent-master/
google-search

A Playwright-based Node.js tool that bypasses search engine anti-scraping mechanisms to execute Google searches. Local alternative to SERP APIs with MCP server integration.

622
usestring/
powhttp-mcp

MCP server enabling agents to debug HTTP requests better (using powhttp)

81
brettdavies/
crawl4ai-skill

Scrape JavaScript-heavy sites and extract structured data via reusable CSS schemas. Portable agent skill wrapping the Crawl4AI CLI and Python SDK.

46

Extract public Spotify data β€” tracks, albums, artists, playlists, podcasts & lyrics β€” without the official API. Sync + async, typed models, one dependency.

301

MCP server for AI agent payments: pay-per-call APIs (search, scraping, data, compute, media, research) with no vendor keys and a spend ceiling on every call. Works in Cursor, Claude Code, Claude Desktop, and Codex.

43

The Apify MCP server enables your AI agents to extract data from social media, search engines, maps, e-commerce sites, or any other website using thousands of ready-made scrapers, crawlers, and automation tools available on the Apify Store.

6.7k
yusufkaraaslan/
Skill_Seekers

Convert documentation websites, GitHub repositories, and PDFs into Claude AI skills with automatic conflict detection

15k

Open-source MCP server for LinkedIn. Give Claude and any MCP-compatible AI agent access to profiles, companies, jobs, and messages.

3.4k
Johell1NS/
browser-search

A skill for AI agents: search the web with SearXNG, browse with Camofox, bypass protections with CloakBrowser. Anti-hallucination by design. Self-hosted, free, unlimited.

517
pinkpixel-dev/
web-scout-mcp

An MCP server providing web search and content extraction capabilities. Integrates DuckDuckGo search functionality and URL content extraction into your MCP environment, enabling AI assistants to search the web and extract webpage content.

133