Bright Data
Search, Crawl and Scrape any site, at scale, without getting blocked
1.0.1Bright Data is a web data platform that provides proxy infrastructure and scraping APIs. This toolkit enables Arcade agents to search the web, scrape pages, and extract structured data from major platforms at scale without being blocked.
Capabilities
- Web scraping: Fetch any URL and receive clean Markdown output, suitable for downstream LLM consumption or content extraction.
- Search engine queries: Run searches against Google, Bing, or Yandex with control over result count, search type (web/images), and country targeting.
- Structured data feeds: Extract typed, schema-consistent records from a wide range of platforms — including Amazon, LinkedIn, Instagram, Facebook, X, YouTube, Zillow, Booking.com, and ZoomInfo — without writing custom parsers.
Secrets
This toolkit requires no OAuth flow but does require two secrets configured in Arcade.
-
BRIGHTDATA_API_KEY— Your Bright Data account's API key, used to authenticate all requests. Obtain it from the Bright Data dashboard under Account Settings → API Token. A paid or trial Bright Data account is required; free-tier access may have limited or no API access. -
BRIGHTDATA_ZONE— The name of a Bright Data proxy zone (e.g., a Web Unlocker or Scraping Browser zone) that the toolkit routes requests through. Create or find existing zones in the Bright Data control panel under Proxies & Scraping Infrastructure. The zone must be active and have sufficient traffic allocated; the required zone type depends on the target sites and tools used.
See the Arcade secrets configuration guide at https://docs.arcade.dev/en/guides/create-tools/tool-basics/create-tool-secrets, or manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MAKE UP LINKS. IF LINKS ARE NEEDED, FIND THEM WITH A WEB SEARCH FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 1 |