View as Markdown
Scrapeless icon

Scrapeless ACTION

Submit Scrape Job

Submit a new web scraping job with specified target URL and extraction rules. See the documentation
  • Action
  • Read only
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's Scrapeless account once, then configure and run Submit Scrape Job from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "scrapeless-submit-scrape-job",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    scrapeless: { authProvisionId: "apn_xxxxxxx" },
    actor: "Actor",
    inputUrl: "Input URL",
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Submit Scrape Job inputs
Property Type Description
actor Actor string
The actor to use for the scrape job. This can be a specific user or a system account.
Required
inputUrl Input URL string
Target URL to scrape. This is the URL of the web page you want to extract data from.
Required
proxyCountry Proxy Country string
The country to route the request through. This can help in bypassing geo-restrictions.
Required
asyncMode Async Mode boolean
Whether to run the scrape job in asynchronous mode. If set to true, the job will be processed in the background.
Required

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
scrapeless-submit-scrape-job
Version
0.0.4
App
Scrapeless
Authentication
API key
Read-only
Yes
Destructive
No
Open world
Yes