Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface
MCP server for web search, scraping, and browser automation
Bright Data MCP connects an agent to public web search, page scraping, structured data extraction, and remote browser automation. It exposes a large tool set for live web work, including search engines, site-specific extractors, package registry lookups, and browser actions like navigate, click, type, and screenshot.
Builders who want their agent to pull live web data, read blocked pages, or automate browser tasks through MCP.
You can give your agent reliable access to current web data instead of relying on brittle scraping code or stale answers.
What it does
Web search
Search Google, Bing, and Yandex and return results as structured data.
Page scraping
Fetch any URL as Markdown or HTML, with bot detection, CAPTCHA solving, and proxy rotation handled for you.
Structured extraction
Pull clean JSON from supported sites like Amazon, LinkedIn, TikTok, YouTube, X, Reddit, Zillow, and more.
Browser automation
Use a remote browser to navigate, click, type, scroll, screenshot, and inspect page text and HTML.
LLM response collection
Ask ChatGPT, Grok, or Perplexity a prompt and get the responses back as structured data.
Package registry data
Look up npm and PyPI package metadata, README content, versions, and dependencies without scraping registries.
Tool grouping
Load only the tool groups you need, such as ecommerce, social, browser, business, finance, research, app stores, travel, geo, code, and advanced scraping.
How to get it
- 1Hosted server — no installation. Add this URL to your MCP client
https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN_HERE
- 2Run
claude mcp add --transport http brightdata "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
README
Bright Data MCP
Web search, page scraping, structured data extraction, and browser automation for AI agents and LLMs over the Model Context Protocol.
Works with AI agents, coding agents, chat assistants, and any MCP-compatible client.
Quick Start • Pricing • Use Cases • Tools • Agent Skills • Docs • Support
Free tier: 5,000 requests per month. No credit card required. Renews monthly.
Overview
The Bright Data MCP server gives AI agents real-time access to public web data. It exposes 69 tools covering:
- Web search — Google, Bing, and Yandex results as structured data
- Page scraping — any URL as Markdown or HTML, with bot detection, CAPTCHA solving, and proxy rotation handled automatically on every request
- Structured data extraction — clean JSON from Amazon, LinkedIn, Instagram, TikTok, YouTube, X, Reddit, Facebook, Crunchbase, Zillow, and other major platforms, without parsing HTML
- Browser automation — navigate, click, type, screenshot, and read pages in a remote browser session
- LLM response collection — send prompts to ChatGPT, Grok, and Perplexity and get their answers back as structured data
- Package registry data — npm and PyPI package versions, READMEs, dependencies, and metadata
Every request is routed through Bright Data's unblocking infrastructure, so pages that block ordinary HTTP clients (bot detection, CAPTCHAs, rate limits, geo-restrictions) return normally. No proxy setup, no headless browser maintenance, no retry logic to write.
Two deployment options: a hosted remote server (one URL, no installation) or a local instance via npx @brightdata/mcp.
Quick Start
Hosted server — no installation. Add this URL to your MCP client:
https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN_HERE
Get your API token from your Bright Data account settings. New accounts get 5,000 free requests per month.
Optional URL parameters:
| Parameter | Description | Example |
|---|---|---|
groups=<ids> | Enable specific tool groups | ...&groups=social,ecommerce |
tools=<names> | Enable specific tools only | ...&tools=search_engine,scrape_as_markdown |
Claude Desktop
- Go to: Settings → Connectors → Add custom connector
- Name:
Bright Data - URL:
https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN - Click "Add"
Or run locally:
{
"mcpServers": {
"Bright Data": {
"command": "npx",
"args": ["@brightdata/mcp"],
"env": {
"API_TOKEN": "<your-api-token-here>"
}
}
}
}
Claude Code
claude mcp add --transport http brightdata "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
Cursor
Add to ~/.cursor/mcp.json:
{
"mcpServers": {
"brightdata": {
"url": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
}
VS Code
Add to .vscode/mcp.json:
{
"servers": {
"brightdata": {
"type": "http",
"url": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
}
Windsurf
Add to ~/.codeium/windsurf/mcp_config.json:
{
"mcpServers": {
"brightdata": {
"serverUrl": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
}
Gemini CLI
Add to ~/.gemini/settings.json:
{
"mcpServers": {
"brightdata": {
"httpUrl": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
}
Zed
Add to your Zed settings:
{
"context_servers": {
"brightdata": {
"url": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
}
Warp
Go to Settings > MCP Servers > Add MCP Server and add:
{
"brightdata": {
"url": "https://mcp.brightdata.com/mcp?token=YOUR_API_TOKEN"
}
}
Other clients (local npx)
For any client that supports local MCP servers:
{
"mcpServers": {
"Bright Data": {
"command": "npx",
"args": ["@brightdata/mcp"],
"env": {
"API_TOKEN": "<your-api-token-here>"
}
}
}
}
Pricing and Free Tier
Every account includes a recurring monthly free tier. No credit card or commitment required to start.
5,000 free requests per month, renewing on the 1st of each month. Unused requests don't roll over. For team accounts, the free tier is shared across all users in the account.
What's included free:
- Fetch any webpage and extract as Markdown
- Access to 60+ pre-built scrapers for popular domains
- Web search (Google, Bing, Yandex)
- Web unlocking (bot detection bypass, CAPTCHA solving, proxy rotation)
- Browser automation
- Geo-targeting
Beyond the free tier — pay as you go, no commitment:
| Search, Scrape & Extract | Browser Navigation | |
|---|---|---|
| Pay as you go | $1.50 / 1K results | $8 / GB |
- When free requests run out, requests stop. No surprise charges — unless you have deposited funds
- Adding a credit card is a verification step only; you are not charged unless your free tier is exhausted and you have funds deposited
- Set a spend cap in the control panel so pay-as-you-go usage never exceeds your budget
Full pricing, volume plans and enterprise →
Use Cases
Real-time research
Answer questions using live web data instead of training data. Search, then read the sources.
| Task | Tools |
|---|---|
| Search the web for current information | search_engine, search_engine_batch |
| Read a specific page as clean Markdown | scrape_as_markdown, scrape_batch |
| Find the most relevant sources for a research question, ranked by AI relevance score | discover |
Example prompts: "What's Tesla's current stock price?", "Get today's weather forecast for New York", "Find the most cited sources on EU AI regulation from the last 6 months".
E-commerce intelligence
Read product data as structured JSON: price, availability, rating, review count, seller, images.
| Task | Tools |
|---|---|
| Amazon product details, reviews, search results | web_data_amazon_product, web_data_amazon_product_reviews, web_data_amazon_product_search |
| Walmart, eBay, Best Buy, Etsy, Home Depot, Zara products | web_data_walmart_product, web_data_ebay_product, web_data_bestbuy_products, web_data_etsy_products, web_data_homedepot_products, web_data_zara_products |
| Cross-retailer price view | web_data_google_shopping |
| Seller profiles | web_data_walmart_seller |
Example prompts: "Compare this laptop's price on Amazon vs Walmart vs Best Buy", "Get the rating and review count for ASIN B0D2Q9397Y", "Is this product in stock?".
Market and competitor analysis
Build competitor profiles from live data: funding, headcount, hiring, customer reviews, pricing pages.
| Task | Tools |
|---|---|
| Company funding, investors, size | web_data_crunchbase_company, web_data_zoominfo_company_profile |
| Company pages, employees, job postings | web_data_linkedin_company_profile, web_data_linkedin_job_listings |
| Customer sentiment | web_data_google_maps_reviews, web_data_facebook_company_reviews, app store review tools |
| Competitor pricing pages | scrape_as_markdown, scrape_batch |
| Market discovery | search_engine_batch, discover |
Example prompt: "Analyze Notion as a competitor: pricing, funding, hiring focus, and what customers complain about".
AI agents with reliable web access
Replace built-in fetch/search tools that get blocked on protected sites. Every request goes through unblocking infrastructure, so agents don't fail on bot detection, CAPTCHAs, or geo-restrictions.
| Task | Tools |
|---|---|
| Drop-in replacement for built-in web search | search_engine |
| Drop-in replacement for built-in URL fetch | scrape_as_markdown |
| Parallel data collection (10 at a time) | search_engine_batch, scrape_batch |
| Interactive sites (login walls, infinite scroll, dynamic content) | scraping_browser_* (13 tools) |
| Structured JSON from any page, no schema needed | extract |
Coding agents
Package registry data on demand — no scraping, no stale caches.
| Task | Tools |
|---|---|
| npm package version, README, dependencies, metadata | web_data_npm_package |
| PyPI package version, README, dependencies, metadata | web_data_pypi_package |
| Read files from GitHub repositories | web_data_github_repository_file |
Example prompts: "What's the latest version of express on npm?", "Get the README for the langchain-brightdata PyPI package".
GEO and brand visibility
Send prompts to major LLMs and get their answers back as structured data. Measure how AI assistants describe your brand, which sources they cite, and what they recommend — the feedback loop for Generative Engine Optimization.
| Task | Tools |
|---|---|
| ChatGPT answers with citations and recommendations | web_data_chatgpt_ai_insights |
| Grok answers | web_data_grok_ai_insights |
| Perplexity answers with sources | web_data_perplexity_ai_insights |
Example prompt: "Ask ChatGPT, Grok, and Perplexity 'what is the best proxy provider' and compare how each one ranks us".
Social media monitoring
Structured data from seven platforms: profiles, posts, comments, engagement metrics.
| Platform | Tools |
|---|---|
| person profiles, company profiles, job listings, posts, people search (5 tools) | |
| profiles, posts, reels, comments (4 tools) | |
| TikTok | profiles, posts, shop, comments (4 tools) |
| posts, marketplace listings, company reviews, events (4 tools) | |
| YouTube | videos, channel profiles, comments (3 tools) |
| X (Twitter) | posts, profile posts (2 tools) |
| posts (1 tool) |
Example prompt: "Get the last 10 posts from this TikTok profile and summarize the engagement".
Content creation and academic research
Gather source material from many pages at once, filtered by recency and relevance.
| Task | Tools |
|---|---|
| Collect multiple sources in one call | scrape_batch (up to 10 URLs) |
| Find sources by topic with date filtering | discover with start_date / end_date |
| News and finance data | web_data_yahoo_finance_business, search_engine with news queries |
How It Compares
| Capability | Bright Data MCP | Typical web MCP servers |
|---|---|---|
| Total tools | 69 | 2–10 |
| Platform-specific structured JSON extractors | 45 tools across e-commerce, social, business, finance, travel, app stores | Rare; generic scraping only |
| Unblocking (bot detection bypass, CAPTCHA solving, proxy rotation) | Built into every request | Usually none; blocked on protected sites |
| Search engines | Google, Bing, Yandex | Usually one |
| AI-relevance-ranked search with intent | Yes (discover) | Not offered |
| Browser automation | 13 tools, remote browser, no local setup | Limited or none |
| LLM response collection (ChatGPT, Grok, Perplexity) | Yes | Not offered |
| Package registry data (npm, PyPI) | Yes | Not offered |
| Batch operations | 10 searches or 10 scrapes per call | Usually single-request only |
| Geo-targeting | Yes | Limited or none |
| Free tier | 5,000 requests/month, browser automation included, no credit card | Varies; often rate-limited keyless access |
Tool Selection: Groups
Tools are organized into groups so you only load what you need. Fewer tools means less context for your agent to process.
GROUPSenables tool bundles. Comma-separated:GROUPS="ecommerce,browser"(local) or&groups=ecommerce,browser(hosted URL)TOOLSadds individual tools on top:TOOLS="extract,scrape_as_html"- Base tools are always enabled:
search_engine,search_engine_batch,scrape_as_markdown,scrape_batch,discover - Group ID
customis reserved; useTOOLSfor individual picks
| Group ID | Contents | Tool count |
|---|---|---|
ecommerce | Amazon, Walmart, eBay, Best Buy, Etsy, Home Depot, Zara, Google Shopping | 11 |
social | LinkedIn, Instagram, Facebook, TikTok, YouTube, X, Reddit | 23 |
browser | Remote browser automation | 13 |
business | Crunchbase, ZoomInfo, Google Maps reviews, Zillow | 4 |
finance | Yahoo Finance | 1 |
research | GitHub repository files | 1 |
app_stores | Google Play, Apple App Store | 2 |
travel | Booking.com | 1 |
geo | ChatGPT, Grok, Perplexity response collection | 3 |
code | npm, PyPI package data | 2 |
advanced_scraping | Batch tools, HTML scraping, AI extraction, session stats | 5 |
Configuration examples
Local server with browser automation and AI extraction:
{
"mcpServers": {
"Bright Data": {
"command": "npx",
"args": ["@brightdata/mcp"],
"env": {
"API_TOKEN": "<your-api-token-here>",
"GROUPS": "browser,advanced_scraping",
"TOOLS": "extract"
}
}
}
}
Coding agent setup (Claude Code / Cursor / Windsurf) — npm and PyPI package data:
{
"mcpServers": {
"Bright Data": {
"command": "npx",
"args": ["@brightdata/mcp"],
"env": {
"API_TOKEN": "<your-api-token-here>",
"GROUPS": "code"
}
}
}
}
Tools Reference (69 Tools)
Which tool to use
- Known URL, need the content:
scrape_as_markdown. Multiple URLs (up to 10):scrape_batch - Need to find information:
search_engine. Multiple queries (up to 10):search_engine_batch - Deep research or RAG, need relevance-ranked sources:
discoverwith anintent - Page is on a supported platform (Amazon, LinkedIn, TikTok, etc.): use the matching
web_data_*tool — returns clean JSON, faster and more reliable than scraping the same page - Structured JSON from an unsupported page:
extract - Raw HTML:
scrape_as_html - Page requires interaction (click, type, scroll, login):
scraping_browser_*tools - npm/PyPI package info:
web_data_npm_package/web_data_pypi_package— never scrape package registries - How ChatGPT/Grok/Perplexity answer a prompt:
web_data_chatgpt_ai_insights/web_data_grok_ai_insights/web_data_perplexity_ai_insights
Notes that apply to all web_data_* tools:
- Return structured JSON, billed per record returned
- Each tool validates its URL pattern; a wrong URL type fails (exact requirements in the tables below)
- Results can be large. Use built-in limits where available (
num_of_comments,days_limit) and run bulk collection in a subagent where your framework supports it, so records don't flood the main context window - If a
web_data_*call fails,scrape_as_markdownworks on the same URL as a fallback
Search and Scraping — 8 tools
| Tool | Description | Group |
|---|---|---|
search_engine | Search Google, Bing, or Yandex. Google returns JSON (URL, title, description); Bing and Yandex return Markdown. Paginate with the cursor parameter | always enabled |
search_engine_batch | Up to 10 search queries in one call | always enabled |
scrape_as_markdown | Any URL as Markdown. Bot protection and CAPTCHA handled automatically | always enabled |
scrape_batch | Up to 10 URLs in one call; returns an array of URL/content pairs in Markdown | always enabled |
discover | AI-relevance-ranked web search. Returns scored results (title, description, URL, relevance score). Supports intent-based ranking, geo-targeting, date filtering, keyword filtering | always enabled |
scrape_as_html | Any URL as raw HTML | advanced_scraping |
extract | Scrape a page and convert it to structured JSON using AI, with an optional custom extraction prompt | advanced_scraping |
session_stats | Tool usage counts for the current session | advanced_scraping |
E-commerce — 11 tools
| Tool | Input requirement | Returns |
|---|---|---|
web_data_amazon_product | Product URL containing /dp/ | Price, title, availability, rating, review count, ASIN, seller, images |
web_data_amazon_product_reviews | Product URL containing /dp/ | Review data |
web_data_amazon_product_search | Search keyword + Amazon domain URL | First page of search results |
web_data_walmart_product | Product URL containing /ip/ | Product data |
web_data_walmart_seller | Walmart seller URL | Seller data |
web_data_ebay_product | eBay product URL | Listing data |
web_data_homedepot_products | homedepot.com product URL | Product data |
web_data_zara_products | Zara product URL | Product data |
web_data_etsy_products | Etsy product URL | Listing data |
web_data_bestbuy_products | Best Buy product URL | Product data |
web_data_google_shopping | Google Shopping product URL | Multi-seller product data |
Social Media — 23 tools
| Tool | Input requirement | Returns |
|---|---|---|
web_data_linkedin_person_profile | LinkedIn profile URL | Profile, experience, skills |
web_data_linkedin_company_profile | LinkedIn company URL | Company data |
web_data_linkedin_job_listings | LinkedIn jobs URL | Job listing data |
web_data_linkedin_posts | LinkedIn post URL | Post data |
web_data_linkedin_people_search | LinkedIn people search URL | Search results |
web_data_instagram_profiles | Instagram profile URL | Profile data |
web_data_instagram_posts | Instagram post URL | Post data |
web_data_instagram_reels | Instagram reel URL | Reel data |
web_data_instagram_comments | Instagram URL | Comments |
web_data_facebook_posts | Facebook post URL | Post data |
web_data_facebook_marketplace_listings | Marketplace listing URL | Listing data |
web_data_facebook_company_reviews | Facebook company URL + review count | Reviews |
web_data_facebook_events | Facebook event URL | Event data |
web_data_tiktok_profiles | TikTok profile URL | Profile data |
web_data_tiktok_posts | TikTok post URL | Post data |
web_data_tiktok_shop | TikTok Shop product URL | Product data |
web_data_tiktok_comments | TikTok video URL | Comments |
web_data_x_posts | X post URL | Post data |
web_data_x_profile_posts | X profile URL | Recent posts, optional date range filter |
web_data_youtube_videos | YouTube video URL | Video metadata |
web_data_youtube_profiles | YouTube channel URL | Channel data |
web_data_youtube_comments | YouTube video URL, optional num_of_comments (default 10) | Comments |
web_data_reddit_posts | Reddit post URL | Post data |
Browser Automation — 13 tools
Remote browser session. Typical sequence: navigate → snapshot → interact by ref → extract or screenshot.
| Tool | Description |
|---|---|
scraping_browser_navigate | Open or reuse a browser session and navigate to a URL |
scraping_browser_go_back | Navigate back |
scraping_browser_go_forward | Navigate forward |
scraping_browser_snapshot | ARIA snapshot of the page listing interactive elements with refs. Required before ref-based actions |
scraping_browser_click_ref | Click an element by ref from the latest snapshot |
scraping_browser_type_ref | Type into an element by ref; optionally press Enter to submit |
scraping_browser_screenshot | Screenshot of the current page; optional full_page |
scraping_browser_get_text | Text content of the page body |
scraping_browser_get_html | HTML of the current page |
scraping_browser_scroll | Scroll to the bottom of the page |
scraping_browser_scroll_to_ref | Scroll an element into view |
scraping_browser_wait_for_ref | Wait for an element to become visible, with optional timeout |
scraping_browser_network_requests | Network requests since page load: method, URL, status |
Refs come from the latest snapshot. If the page changes after a click or navigation, take a new snapshot before the next ref-based action. For static pages, scrape_as_markdown is faster and cheaper than a browser session.
Business Intelligence — 4 tools
| Tool | Input requirement | Returns |
|---|---|---|
web_data_crunchbase_company | Crunchbase company URL | Funding, investors, company data |
web_data_zoominfo_company_profile | ZoomInfo company URL | Company profile |
web_data_google_maps_reviews | Google Maps URL, optional days_limit (default 3) | Business reviews |
web_data_zillow_properties_listing | Zillow listing URL | Property listing data |
GEO and LLM Visibility — 3 tools
| Tool | Input | Returns |
|---|---|---|
web_data_chatgpt_ai_insights | Prompt | ChatGPT's answer: structured text, citations, recommendations, Markdown |
web_data_grok_ai_insights | Prompt | Grok's answer as structured Markdown |
web_data_perplexity_ai_insights | Prompt | Perplexity's answer with sources, as structured Markdown |
Use for Generative Engine Optimization (tracking how LLMs describe your brand) and LLM-as-a-judge workflows.
Code — 2 tools
| Tool | Input | Returns |
|---|---|---|
web_data_npm_package | npm package name (e.g., @brightdata/sdk) | Latest version, README, dependencies, metadata |
web_data_pypi_package | PyPI package name (e.g., langchain-brightdata) | Latest version, README, dependencies, metadata |
Finance, Research, App Stores, Travel — 5 tools
| Tool | Input requirement | Returns | Group |
|---|---|---|---|
web_data_yahoo_finance_business | Yahoo Finance business URL | Company financial data | finance |
web_data_github_repository_file | GitHub file URL | File content and metadata | research |
web_data_google_play_store | Play Store app URL | App details | app_stores |
web_data_apple_app_store | App Store app URL | App details | app_stores |
web_data_booking_hotel_listings | Booking.com listing URL | Hotel listing data | travel |
Full tool reference in the docs →
Agent Skills
Ready-to-use skills that teach your agent how to use this MCP server correctly. The full
Files in the repo
- .github
- assets
- examples
- mcp-evals
- test
- .gitignore
- .mcpbignore
- .npmignore
- aria_snapshot_filter.js
- brightdata-mcp-2.7.1.mcpb
- browser_session.js
- browser_tools.js
- CHANGELOG.md
- Dockerfile
- icon.png
- LICENSE
- manifest.json
- package-lock.json
- package.json
- prompts.js
- README.md
- search_dataset_schema.js
- search_utils.js
- server.js
- server.json
- smithery.yaml
- tool_groups.js
Discussion (0)
Ask about usage, or say what you built with itSign in to join the discussion.
No comments yet. Be the first to say what this is good for.
More connectors
High-performance code intelligence MCP server. Indexes codebases into a persistent knowledge graph — average repo in milliseconds. 158 languages, sub-ms queries, 99% fewer tokens. Single static binary, zero dependencies.

Universal provider proxy for OpenAI Codex & Claude Code — use any LLM (Claude, Gemini, Grok, DeepSeek, Ollama…) with Codex CLI, App, SDK, and Claude Code
Git-native persistent memory for AI coding agents. Implements Google OKF v0.2 with sub-300µs in-memory BM25 search, embedded MCP server, and progressive disclosure. Slashes token bloat by 80% with zero external databases or dependencies. Built in pure Go.
Local-first code intelligence graph for MCP and CLI. Builds a persistent map of your codebase so AI coding tools read only what matters, with benchmarked context reductions on reviews and large-repo workflows.
Stop your AI from making things up — it proposes, deterministic tools decide, every claim checked against ground truth with evidence. Grounded facts and context survive resets. Reverse engineering is the proving ground. MCP server + CLI.