Reference for the Bright Data MCP server tools (60+ tools): web search, scraping, structured data and platform extractors for ecommerce and social workflows.
Building an AI startup?
You might be eligible for our Startup Program. Get fully funded access to the infrastructure you’re reading about right now (up to $20K value).
Bright Data’s MCP server offers two modes to suit different needs:
Rapid (Free) - Quickly scrape search results and unlock any public webpage as clean Markdown.
Pro - Access advanced scraping, structured data from top platforms (Amazon, LinkedIn, X, Instagram, etc.), and full browser automation. Built for dynamic and large-scale use cases.
Groups - Pre-configured tool collections for specific use cases including E-commerce, Social Media, Browser Automation, Business Intelligence, Finance, Research, App Stores, Travel, Advanced Scraping, GEO & LLM Visibility, and Code. Each group bundles relevant tools and data feeds for its domain.
To use Pro mode:
In Remote MCP, request Pro features by setting &pro=1.
In Local MCP, enable Pro features with PRO_MODE=true.
To use Groups:
In Remote MCP, request the Groups by setting &groups=<group_name>.
In Local MCP, enable Groups with GROUPS=<group_name>.
Rapid (Free) mode is enabled by default and is recommended for everyday browsing and data needs.
Scrape search results from Google, Bing, or Yandex. Returns SERP results in JSON for Google and Markdown for Bing/Yandex; supports pagination with the cursor parameter.
all
scrape_as_markdown
Scrape a single webpage with advanced extraction and return Markdown. Uses Bright Data’s unlocker to handle bot protection and CAPTCHA.
all
discover
Search the web and rank results by AI-driven relevance. Returns scored results with title, description, and URL. Supports intent-based ranking, geo-targeting, date filtering, and keyword filtering.
all
search_engine_batch
Run up to 10 search queries in parallel. Returns JSON for Google results and Markdown for Bing/Yandex.
advanced_scraping
scrape_batch
Scrape up to 10 webpages in one request and return an array of URL/content pairs in Markdown format.
advanced_scraping
scrape_as_html
Scrape a single webpage with advanced extraction and return the HTML response body. Handles sites protected by bot detection or CAPTCHA.
advanced_scraping
extract
Scrape a webpage as Markdown and convert it to structured JSON using AI sampling, with an optional custom extraction prompt.
advanced_scraping
session_stats
Report how many times each tool has been called during the current MCP session.
advanced_scraping
web_data_amazon_product
Quickly read structured Amazon product data. Requires a valid product URL containing /dp/. Often faster and more reliable than scraping.
ecommerce
web_data_amazon_product_reviews
Quickly read structured Amazon product review data. Requires a valid product URL containing /dp/. Often faster and more reliable than scraping.
ecommerce
web_data_amazon_product_search
Retrieve structured Amazon search results. Requires a search keyword and Amazon domain URL; limited to the first page of results.
ecommerce
web_data_walmart_product
Quickly read structured Walmart product data. Requires a product URL containing /ip/. Often faster and more reliable than scraping.
ecommerce
web_data_walmart_seller
Quickly read structured Walmart seller data. Requires a valid Walmart seller URL. Often faster and more reliable than scraping.
ecommerce
web_data_ebay_product
Quickly read structured eBay product data. Requires a valid eBay product URL. Often faster and more reliable than scraping.
ecommerce
web_data_homedepot_products
Quickly read structured Home Depot product data. Requires a valid homedepot.com product URL. Often faster and more reliable than scraping.
ecommerce
web_data_zara_products
Quickly read structured Zara product data. Requires a valid Zara product URL. Often faster and more reliable than scraping.
ecommerce
web_data_etsy_products
Quickly read structured Etsy product data. Requires a valid Etsy product URL. Often faster and more reliable than scraping.
ecommerce
web_data_bestbuy_products
Quickly read structured Best Buy product data. Requires a valid Best Buy product URL. Often faster and more reliable than scraping.
ecommerce
web_data_linkedin_person_profile
Quickly read structured LinkedIn people profile data. Requires a valid LinkedIn profile URL. Often faster and more reliable than scraping.
social
web_data_linkedin_company_profile
Quickly read structured LinkedIn company profile data. Requires a valid LinkedIn company URL. Often faster and more reliable than scraping.
social
web_data_linkedin_job_listings
Quickly read structured LinkedIn job listings data. Requires a valid LinkedIn jobs URL. Often faster and more reliable than scraping.
social
web_data_linkedin_posts
Quickly read structured LinkedIn posts data. Requires a valid LinkedIn post URL. Often faster and more reliable than scraping.
social
web_data_linkedin_people_search
Quickly read structured LinkedIn people search data. Requires a LinkedIn people search URL. Often faster and more reliable than scraping.
social
list_dataset_fields
List the filterable fields of a searchable dataset with each field’s name, type and description. Call before search_dataset to learn which fields you can filter on.
social, business
search_dataset
Search a Bright Data dataset by a filter and get matching records back directly, with no snapshot step. Finds many records by criteria, unlike web_data_* tools that fetch one record by URL. Returns up to 10 records per call with search_after pagination.
social, business
web_data_crunchbase_company
Quickly read structured Crunchbase company data. Requires a valid Crunchbase company URL. Often faster and more reliable than scraping.
business
web_data_zoominfo_company_profile
Quickly read structured ZoomInfo company profile data. Requires a valid ZoomInfo company URL. Often faster and more reliable than scraping.
business
web_data_instagram_profiles
Quickly read structured Instagram profile data. Requires a valid Instagram profile URL. Often faster and more reliable than scraping.
social
web_data_instagram_posts
Quickly read structured Instagram post data. Requires a valid Instagram post URL. Often faster and more reliable than scraping.
social
web_data_instagram_reels
Quickly read structured Instagram reel data. Requires a valid Instagram reel URL. Often faster and more reliable than scraping.
social
web_data_instagram_comments
Quickly read structured Instagram comments data. Requires a valid Instagram URL. Often faster and more reliable than scraping.
social
web_data_facebook_posts
Quickly read structured Facebook post data. Requires a valid Facebook post URL. Often faster and more reliable than scraping.
social
web_data_facebook_marketplace_listings
Quickly read structured Facebook Marketplace listing data. Requires a valid Marketplace listing URL. Often faster and more reliable than scraping.
social
web_data_facebook_company_reviews
Quickly read structured Facebook company reviews data. Requires a valid Facebook company URL and review count. Often faster and more reliable than scraping.
social
web_data_facebook_events
Quickly read structured Facebook events data. Requires a valid Facebook event URL. Often faster and more reliable than scraping.
social
web_data_tiktok_profiles
Quickly read structured TikTok profile data. Requires a valid TikTok profile URL. Often faster and more reliable than scraping.
social
web_data_tiktok_posts
Quickly read structured TikTok post data. Requires a valid TikTok post URL. Often faster and more reliable than scraping.
social
web_data_tiktok_shop
Quickly read structured TikTok Shop product data. Requires a valid TikTok Shop product URL. Often faster and more reliable than scraping.
social
web_data_tiktok_comments
Quickly read structured TikTok comments data. Requires a valid TikTok video URL. Often faster and more reliable than scraping.
social
web_data_google_maps_reviews
Quickly read structured Google Maps reviews data. Requires a valid Google Maps URL and optional days_limit (default 3). Often faster and more reliable than scraping.
business
web_data_google_shopping
Quickly read structured Google Shopping product data. Requires a valid Google Shopping product URL. Often faster and more reliable than scraping.
ecommerce
web_data_google_play_store
Quickly read structured Google Play Store app data. Requires a valid Play Store app URL. Often faster and more reliable than scraping.
app_stores
web_data_apple_app_store
Quickly read structured Apple App Store app data. Requires a valid App Store app URL. Often faster and more reliable than scraping.
app_stores
web_data_github_repository_file
Quickly read structured GitHub repository file data. Requires a valid GitHub file URL. Often faster and more reliable than scraping.
research
web_data_yahoo_finance_business
Quickly read structured Yahoo Finance company profile data. Requires a valid Yahoo Finance business URL. Often faster and more reliable than scraping.
finance
web_data_x_posts
Quickly read structured X (Twitter) post data. Requires a valid X post URL. Often faster and more reliable than scraping.
social
web_data_x_profile_posts
Quickly read structured X posts from a profile. Requires a valid X profile URL. Returns the most recent posts, with optional date range filtering.
social
web_data_zillow_properties_listing
Quickly read structured Zillow property listing data. Requires a valid Zillow listing URL. Often faster and more reliable than scraping.
business
web_data_booking_hotel_listings
Quickly read structured Booking.com hotel listing data. Requires a valid Booking.com listing URL. Often faster and more reliable than scraping.
travel
web_data_youtube_videos
Quickly read structured YouTube video metadata. Requires a valid YouTube video URL. Often faster and more reliable than scraping.
social
web_data_youtube_profiles
Quickly read structured YouTube channel profile data. Requires a valid YouTube channel URL. Often faster and more reliable than scraping.
social
web_data_youtube_comments
Quickly read structured YouTube comments data. Requires a valid YouTube video URL and optional num_of_comments (default 10). Often faster and more reliable than scraping.
social
web_data_reddit_posts
Quickly read structured Reddit post data. Requires a valid Reddit post URL. Often faster and more reliable than scraping.
social
scraping_browser_navigate
Open or reuse a scraping-browser session and navigate to the provided URL, resetting tracked network requests.
browser
scraping_browser_go_back
Navigate the active scraping-browser session back to the previous page and report the new URL and title.
browser
scraping_browser_go_forward
Navigate the active scraping-browser session forward to the next page and report the new URL and title.
browser
scraping_browser_snapshot
Capture an ARIA snapshot of the current page listing interactive elements and their refs for later ref-based actions.
browser
scraping_browser_click_ref
Click an element using its ref from the latest ARIA snapshot; requires a ref and human-readable element description.
browser
scraping_browser_type_ref
Fill an element identified by ref from the ARIA snapshot, optionally pressing Enter to submit after typing.
browser
scraping_browser_screenshot
Capture a screenshot of the current page; supports optional full_page mode for full-length images.
browser
scraping_browser_network_requests
List the network requests recorded since page load with HTTP method, URL, and response status for debugging.
browser
scraping_browser_wait_for_ref
Wait until an element identified by ARIA ref becomes visible, with an optional timeout in milliseconds.
browser
scraping_browser_get_text
Return the text content of the current page’s body element.
browser
scraping_browser_get_html
Return the HTML content of the current page; avoid the full_page option unless head or script tags are required.
browser
scraping_browser_scroll
Scroll to the bottom of the current page in the scraping-browser session.
browser
scraping_browser_scroll_to_ref
Scroll the page until the element referenced in the ARIA snapshot is in view.
browser
web_data_chatgpt_ai_insights
Send a prompt to ChatGPT and get back AI-generated insights. Returns structured answer text, citations, recommendations, and markdown. Useful for GEO and LLM as a judge.
geo
web_data_grok_ai_insights
Send a prompt to Grok and get back AI-generated insights. Returns structured answer text in markdown format. Useful for GEO and LLM as a judge.
geo
web_data_perplexity_ai_insights
Send a prompt to Perplexity and get back AI-generated insights. Returns structured answer text in markdown format. Useful for GEO and LLM as a judge.
geo
web_data_npm_package
Quickly read structured npm package data. Requires a valid npm package name (e.g., @brightdata/sdk). Often faster and more reliable than scraping.
code
web_data_pypi_package
Quickly read structured PyPI package data. Requires a valid PyPI package name (e.g., langchain-brightdata). Often faster and more reliable than scraping.
How do I search datasets by filter with the Bright Data MCP server?
Use list_dataset_fields and search_dataset to find many records that match your criteria, as opposed to the web_data_* tools, which fetch one record by URL. The search runs directly against the dataset with no snapshot step and returns up to 10 records per call. Under the hood, list_dataset_fields calls the Get dataset metadata endpoint and search_dataset calls the Search dataset endpoint of the Marketplace Dataset API.Three datasets are searchable:
dataset_id
Dataset
gd_l1viktl72bvl7bjuj0
LinkedIn people profiles
gd_me5ppxjr2ge6icjuh0
LinkedIn people profiles (contact-enriched)
gd_l1vikfnt1wgvvqz95w
LinkedIn company information
The workflow is two calls:
Call list_dataset_fields with a dataset_id. It returns each filterable field’s name, type and description.
Call search_dataset with the same dataset_id and a filter.
A filter is a tree with a maximum nesting depth of three. Each node is either a group, {"operator": "and" | "or", "filters": [...]}, or a leaf, {"name": ..., "operator": ..., "value": ...}. See filter syntax for the full list of leaf operators and their behavior.For example, this filter matches profiles in the United States whose position mentions data engineering:
The response contains hits (the matching records), total_hits and took (search time in milliseconds). To page through more than 10 results, set sort to "default" or a custom sort like [{"timestamp": "asc"}], then pass the response’s search_after cursor into the next call. Set sort to "random" to sample records instead.