Bright Data
Search, Crawl and Scrape any site, at scale, without getting blocked
1.0.1Bright Data is a web data platform; this toolkit enables Arcade tools to scrape, search, and extract structured data from any public website at scale without being blocked.
Capabilities
- Web scraping: Fetch any webpage and return its content as clean Markdown, suitable for LLM consumption or downstream processing.
- Search engine queries: Run searches against Google, Bing, or Yandex with configurable parameters including result count, country code, and content type (web, images, etc.).
- Structured data extraction: Pull pre-parsed, schema'd data from major platforms — Amazon products and reviews, LinkedIn people and companies, Instagram profiles/posts/reels/comments, Facebook posts/marketplace/reviews, X posts, Zillow listings, Booking.com hotels, YouTube videos, and ZoomInfo company profiles.
Secrets
This toolkit requires no OAuth flow but does require two secrets configured in Arcade.
BRIGHTDATA_API_KEY
Your Bright Data account API key, used to authenticate all API requests. Obtain it from the Bright Data dashboard under API Tokens (or Account Settings → API Token). A paid or trial Bright Data account is required; free-tier access may be limited.
BRIGHTDATA_ZONE
The name of the Bright Data zone (proxy/scraping zone) to route requests through. Zones are created and managed in the Bright Data control panel under My Zones. The zone name is the string identifier you assign when creating a zone (e.g., residential_1 or scraping_browser). The correct zone type must be enabled for the operations you intend to perform (e.g., a Web Unlocker or Scraping Browser zone for scraping, a SERP API zone for search).
For general guidance on configuring secrets in Arcade, see the Arcade secrets docs. You can also manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MAKE UP LINKS. IF LINKS ARE NEEDED, FIND THEM WITH A WEB SEARCH FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 1 |