Skip to Content
ResourcesIntegrationsDeveloper ToolsBright Data

Bright Data

Service domainWEB SCRAPING
Bright Data icon
CommunityBYOC

Search, Crawl and Scrape any site, at scale, without getting blocked

Author:Arcade
Version:1.0.1
Auth:No authentication required
3tools
3require secrets

Bright Data is a web data platform that enables large-scale scraping, searching, and structured data extraction without getting blocked. This toolkit exposes Bright Data's proxy and scraping infrastructure through Arcade tools.

Capabilities

  • Web scraping: Fetch any public webpage and return its content as clean Markdown, suitable for LLM consumption or downstream processing.
  • Multi-engine search: Query Google, Bing, or Yandex with control over result count, search type (web/images), and country targeting.
  • Structured data feeds: Extract pre-parsed, schema'd data from major platforms — Amazon (products, reviews), LinkedIn (people, companies), Instagram, Facebook, X, YouTube, Zillow, Booking.com, and ZoomInfo — without writing custom parsers.

Secrets

This toolkit requires no OAuth flow, but two secrets must be configured before any tool can be called.

  • BRIGHTDATA_API_KEY — Your Bright Data API key, used to authenticate all requests. Obtain it from the Bright Data control panel under Account Settings → API Token. Any paid Bright Data account can generate one; free-tier access may be limited.

  • BRIGHTDATA_ZONE — The name of a Bright Data proxy zone (also called a "zone" or "dataset zone") that the toolkit routes requests through. Create or view zones in the Bright Data control panel. The zone must be active and have sufficient traffic/quota for your use case; different zone types (residential, datacenter, ISP) affect success rates and cost.

Configure both secrets via the Arcade secrets docs or directly in the Arcade dashboard.

Available tools(3)

3 of 3 tools
Operations
Behavior
Tool nameDescriptionSecrets
Scrape a webpage and return content in Markdown format using Bright Data. Examples: scrape_as_markdown("https://example.com") -> "# Example Page Content..." scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News ..."
2
Search using Google, Bing, or Yandex with advanced parameters using Bright Data. Examples: search_engine("climate change") -> "# Search Results ## Climate Change - Wikipedia ..." search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results ..." search_engine("cats", search_type="images", country_code="us") -> "# Image Results ..."
2
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc. NEVER MAKE UP LINKS. IF LINKS ARE NEEDED, FIND THEM WITH A WEB SEARCH FIRST. Supported source types: - amazon_product, amazon_product_reviews - linkedin_person_profile, linkedin_company_profile - zoominfo_company_profile - instagram_profiles, instagram_posts, instagram_reels, instagram_comments - facebook_posts, facebook_marketplace_listings, facebook_company_reviews - x_posts - zillow_properties_listing - booking_hotel_listings - youtube_videos Examples: web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW") -> "{"title": "Product Name", ...}" web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe") -> "{"name": "John Doe", ...}" web_data_feed( "facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50 ) -> "[{"review": "...", ...}]"
1
Last updated on