View as Markdown
Bright Data icon

Bright Data ACTION

Scrape Website

Scrape a website and return the HTML. See the documentation
  • Action
  • Writes data
  • Destructive
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's Bright Data account once, then configure and run Scrape Website from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "bright_data-scrape-website",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    bright_data: { authProvisionId: "apn_xxxxxxx" },
    datasetId: "Dataset ID",
    url: "URL",
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Scrape Website inputs
Property Type Description
datasetId Dataset ID string
The ID of the dataset to use
Required Dynamic
url URL string
The URL of the website to scrape
Required

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
bright_data-scrape-website
Version
0.0.2
App
Bright Data
Authentication
API key
Read-only
No
Destructive
Yes
Open world
Yes