Bright Data
Search, Crawl and Scrape any site, at scale, without getting blocked
1.0.1Bright Data is a web data platform; this toolkit enables Arcade tools to search, scrape, and extract structured data from any site at scale without getting blocked.
Capabilities
- Web scraping: Fetch any public webpage and receive clean Markdown output, suitable for downstream LLM processing or content extraction.
- Search engine queries: Run searches against Google, Bing, or Yandex with control over result count, search type (web/images), and country targeting.
- Structured data feeds: Extract pre-parsed, schema'd data from major platforms — Amazon (products, reviews), LinkedIn (people, companies), Instagram (profiles, posts, reels, comments), Facebook (posts, marketplace, reviews), X posts, Zillow listings, Booking.com hotels, YouTube videos, and ZoomInfo company profiles.
Secrets
BRIGHTDATA_API_KEY — Your Bright Data account API key, used to authenticate all requests to the Bright Data platform. Obtain it from the Bright Data control panel under Account Settings → API Token. A paid Bright Data account is required; free trials may have restricted access to certain products.
BRIGHTDATA_ZONE — The name of the Bright Data proxy/scraping zone your requests should route through. Zones are created and managed in the Bright Data control panel under Proxies & Scraping Infrastructure. Each zone corresponds to a specific product (e.g., Web Unlocker, Scraping Browser, Residential Proxies); choose or create the zone that matches your intended use case and copy its exact name as shown in the dashboard.
For help configuring secrets in Arcade, see the Arcade secrets guide. You can also manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MAKE UP LINKS. IF LINKS ARE NEEDED, FIND THEM WITH A WEB SEARCH FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 1 |