Skip to Content
ResourcesIntegrationsDeveloper ToolsBright Data

Bright Data

Service domainWEB SCRAPING
Bright Data icon
CommunityBYOC

Search, Crawl and Scrape any site, at scale, without getting blocked

Author:Arcade
Version:0.5.1
Auth:No authentication required
3tools
3require secrets

Bright Data is a web data platform that provides infrastructure for scraping, searching, and extracting structured data at scale without getting blocked. The Arcade toolkit exposes Bright Data's capabilities as callable tools for web content retrieval, search, and structured data extraction.

Capabilities

  • Web scraping: Fetch any publicly accessible webpage and return its content as clean Markdown, suitable for downstream processing or LLM consumption.
  • Multi-engine search: Query Google, Bing, or Yandex with configurable parameters including result count, country code, and search type (web or images).
  • Structured data feeds: Extract pre-parsed, schema'd data from major platforms — including Amazon (products, reviews), LinkedIn (people, companies), Instagram (profiles, posts, reels, comments), Facebook (posts, marketplace, reviews), X (posts), Zillow (listings), Booking.com (hotels), YouTube (videos), and ZoomInfo (companies).

Secrets

  • BRIGHTDATA_API_KEY — Your Bright Data account API key, used to authenticate all requests. Obtain it from the Bright Data control panel under Account Settings → API Token. A paid Bright Data account is required; free trials may have restricted access.
  • BRIGHTDATA_ZONE — The Bright Data proxy zone (also called a "dataset" or "zone" identifier) that routes requests. Zones are created and managed in the Bright Data control panel under Proxies & Scraping Infrastructure. The correct zone type depends on your use case (e.g., a Scraping Browser zone for ScrapeAsMarkdown, a Web Unlocker zone for SearchEngine, or a Dataset API zone for WebDataFeed). Copy the zone name exactly as shown in the dashboard.

For configuring secrets in Arcade, see the tool secrets guide. You can also manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.

Available tools(3)

3 of 3 tools
Operations
Behavior
Tool nameDescriptionSecrets
Scrape a webpage and return content in Markdown format using Bright Data. Examples: scrape_as_markdown("https://example.com") -> "# Example Page Content..." scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News ..."
2
Search using Google, Bing, or Yandex with advanced parameters using Bright Data. Examples: search_engine("climate change") -> "# Search Results ## Climate Change - Wikipedia ..." search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results ..." search_engine("cats", search_type="images", country_code="us") -> "# Image Results ..."
2
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc. NEVER MADE UP LINKS - IF LINKS ARE NEEDED, EXECUTE search_engine FIRST. Supported source types: - amazon_product, amazon_product_reviews - linkedin_person_profile, linkedin_company_profile - zoominfo_company_profile - instagram_profiles, instagram_posts, instagram_reels, instagram_comments - facebook_posts, facebook_marketplace_listings, facebook_company_reviews - x_posts - zillow_properties_listing - booking_hotel_listings - youtube_videos Examples: web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW") -> "{"title": "Product Name", ...}" web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe") -> "{"name": "John Doe", ...}" web_data_feed( "facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50 ) -> "[{"review": "...", ...}]"
2
Last updated on