Bright Data
Search, Crawl and Scrape any site, at scale, without getting blocked
0.5.1Bright Data is a web data platform that enables large-scale scraping, searching, and structured data extraction without getting blocked. The Arcade toolkit wraps Bright Data's infrastructure to let developers fetch web content, run search engine queries, and pull structured records from major platforms programmatically.
Capabilities
- Web scraping: Fetch any URL and receive clean Markdown output, handling bot-detection evasion transparently.
- Multi-engine search: Query Google, Bing, or Yandex with configurable result count, search type (web/images), and country targeting.
- Structured data feeds: Extract typed JSON records from 20+ source types across Amazon, LinkedIn, Instagram, Facebook, X, YouTube, Zillow, Booking.com, and ZoomInfo — no parsing required.
- Scale without blocking: All requests route through Bright Data's proxy and unlocker infrastructure, so rate-limiting and bot-detection are handled at the network layer.
Secrets
-
BRIGHTDATA_API_KEY— Your Bright Data API key, used to authenticate all requests. Obtain it from the Bright Data control panel under Account Settings → API Token. A paid Bright Data account is required; free-trial access may have restricted zone and product availability. -
BRIGHTDATA_ZONE— The name of the Bright Data zone (proxy zone or Web Unlocker zone) that requests are routed through. Zones are created and managed in the Bright Data control panel under Proxies & Scraping Infrastructure. The zone name must match exactly (e.g.,web_unlocker1). Different zones carry different pricing and feature sets (datacenter, residential, Web Unlocker, SERP API), so choose the zone type appropriate for your use case before configuring this secret.
For guidance on storing and referencing secrets in Arcade, see the Arcade secrets docs. You can manage your configured secrets at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MADE UP LINKS - IF LINKS ARE NEEDED, EXECUTE search_engine FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 2 |