Bright Data
Service domainWEB SCRAPING
CommunityBYOC
Search, Crawl and Scrape any site, at scale, without getting blocked
Author:Arcade
Version:
0.5.1Auth:No authentication required
3tools
3require secrets
Bright Data is a web data platform that provides infrastructure for scraping, searching, and extracting structured data at scale without getting blocked. The Arcade toolkit exposes Bright Data's capabilities as callable tools for web content retrieval, search, and structured data extraction.
Capabilities
- Web scraping: Fetch any publicly accessible webpage and return its content as clean Markdown, suitable for downstream processing or LLM consumption.
- Multi-engine search: Query Google, Bing, or Yandex with configurable parameters including result count, country code, and search type (web or images).
- Structured data feeds: Extract pre-parsed, schema'd data from major platforms — including Amazon (products, reviews), LinkedIn (people, companies), Instagram (profiles, posts, reels, comments), Facebook (posts, marketplace, reviews), X (posts), Zillow (listings), Booking.com (hotels), YouTube (videos), and ZoomInfo (companies).
Secrets
BRIGHTDATA_API_KEY— Your Bright Data account API key, used to authenticate all requests. Obtain it from the Bright Data control panel under Account Settings → API Token. A paid Bright Data account is required; free trials may have restricted access.BRIGHTDATA_ZONE— The Bright Data proxy zone (also called a "dataset" or "zone" identifier) that routes requests. Zones are created and managed in the Bright Data control panel under Proxies & Scraping Infrastructure. The correct zone type depends on your use case (e.g., a Scraping Browser zone forScrapeAsMarkdown, a Web Unlocker zone forSearchEngine, or a Dataset API zone forWebDataFeed). Copy the zone name exactly as shown in the dashboard.
For configuring secrets in Arcade, see the tool secrets guide. You can also manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
3 of 3 tools
Operations
Behavior
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MADE UP LINKS - IF LINKS ARE NEEDED, EXECUTE search_engine FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 2 |
Last updated on