Bright Data
Search, Crawl and Scrape any site, at scale, without getting blocked
0.5.1Bright Data Toolkit
Bright Data is a web data platform that enables large-scale scraping, searching, and structured data extraction without blocks. This toolkit gives Arcade agents access to Bright Data's proxy and dataset infrastructure for web automation workflows.
Capabilities
- Web scraping: Fetch any publicly accessible webpage and return its content as clean Markdown, suitable for downstream parsing or LLM consumption.
- Multi-engine search: Query Google, Bing, or Yandex with configurable parameters including result count, country code, and search type (web, images, etc.).
- Structured data feeds: Extract pre-structured JSON data from 20+ named source types across major platforms — including Amazon products and reviews, LinkedIn profiles, Instagram content, Facebook posts and reviews, X posts, Zillow listings, Booking.com hotels, YouTube videos, and ZoomInfo company profiles.
Secrets
This toolkit requires two secrets to be configured in Arcade.
-
BRIGHTDATA_API_KEY— Your Bright Data account API key, used to authenticate all requests to Bright Data's APIs. Obtain it from the Bright Data control panel under Account Settings → API Tokens. A paid or trial Bright Data account is required. The key must have permissions to access the proxy zones and datasets your use case requires. -
BRIGHTDATA_ZONE— The name of the Bright Data proxy zone to route scraping and search requests through. Zones are created and managed in the Bright Data control panel under Proxies & Infrastructure → Zones. Create or select a zone (e.g., a residential or datacenter proxy zone), then use its exact name (as shown in the dashboard) as the value for this secret.
For instructions on configuring secrets in Arcade, see the Arcade secrets guide. You can also manage secrets directly at https://api.arcade.dev/dashboard/auth/secrets.
Available tools(3)
| Tool name | Description | Secrets | |
|---|---|---|---|
Scrape a webpage and return content in Markdown format using Bright Data.
Examples:
scrape_as_markdown("https://example.com") -> "# Example Page
Content..."
scrape_as_markdown("https://news.ycombinator.com") -> "# Hacker News
..."
| 2 | ||
Search using Google, Bing, or Yandex with advanced parameters using Bright Data.
Examples:
search_engine("climate change") -> "# Search Results
## Climate Change - Wikipedia
..."
search_engine("Python tutorials", engine="bing", num_results=5) -> "# Bing Results
..."
search_engine("cats", search_type="images", country_code="us") -> "# Image Results
..."
| 2 | ||
Extract structured data from various websites like LinkedIn, Amazon, Instagram, etc.
NEVER MADE UP LINKS - IF LINKS ARE NEEDED, EXECUTE search_engine FIRST.
Supported source types:
- amazon_product, amazon_product_reviews
- linkedin_person_profile, linkedin_company_profile
- zoominfo_company_profile
- instagram_profiles, instagram_posts, instagram_reels, instagram_comments
- facebook_posts, facebook_marketplace_listings, facebook_company_reviews
- x_posts
- zillow_properties_listing
- booking_hotel_listings
- youtube_videos
Examples:
web_data_feed("amazon_product", "https://amazon.com/dp/B08N5WRWNW")
-> "{"title": "Product Name", ...}"
web_data_feed("linkedin_person_profile", "https://linkedin.com/in/johndoe")
-> "{"name": "John Doe", ...}"
web_data_feed(
"facebook_company_reviews", "https://facebook.com/company", num_of_reviews=50
) -> "[{"review": "...", ...}]" | 2 |