View as Markdown
Spider icon

Spider ACTION

Scrape New Page

Initiates a new page scrape (crawl). See the documentation
  • Action
  • Writes data
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's Spider account once, then configure and run Scrape New Page from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "spider-scrape-new-page",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    spider: { authProvisionId: "apn_xxxxxxx" },
    url: "URL",
    limit: 10,
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Scrape New Page inputs
Property Type Description
url URL string
The URI resource to crawl, e.g. https://spider.cloud. This can be a comma split list for multiple urls.
Required
limit Limit integer
The maximum amount of pages allowed to crawl per website. Default is 0, which crawls all pages.
Optional
storeData Store Data boolean
Decide whether to store data. Default is false.
Optional

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
spider-scrape-new-page
Version
0.0.3
App
Spider
Authentication
API key
Read-only
No
Destructive
No
Open world
Yes