View as Markdown
WebScraper.IO icon

WebScraper.IO ACTION

Create Scraping Job

Creates a scraping job (scrapes a sitemap). See the docs here
  • Action
  • Writes data
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's WebScraper.IO account once, then configure and run Create Scraping Job from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "webscraper_io-create-scraping-job",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    webscraper_io: { authProvisionId: "apn_xxxxxxx" },
    sitemapId: "Sitemap ID",
    driver: "Driver",
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Create Scraping Job inputs
Property Type Description
sitemapId Sitemap ID string
Identifier of a sitemap
Required Dynamic
driver Driver string
Driver to use for the scraping job
Optional

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
webscraper_io-create-scraping-job
Version
0.0.2
App
WebScraper.IO
Authentication
API key
Read-only
No
Destructive
No
Open world
Yes