View as Markdown
WebScraper.IO icon

WebScraper.IO ACTION

Get Scraping Jobs

Retrieves a list of scraping jobs for a sitemap. See the docs here
  • Action
  • Read only
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's WebScraper.IO account once, then configure and run Get Scraping Jobs from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "webscraper_io-get-scraping-jobs",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    webscraper_io: { authProvisionId: "apn_xxxxxxx" },
    sitemapId: "Sitemap ID",
    maxResults: 10,
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Get Scraping Jobs inputs
Property Type Description
sitemapId Sitemap ID string
Identifier of a sitemap
Required Dynamic
maxResults Max Results integer
The maximum number of jobs to return
Optional

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
webscraper_io-get-scraping-jobs
Version
0.0.2
App
WebScraper.IO
Authentication
API key
Read-only
Yes
Destructive
No
Open world
Yes