# Scrape New Page — Spider

> Initiates a new page scrape (crawl). See the documentation

- Key: `spider-scrape-new-page`
- Type: Action (Write)
- Version: 0.0.3
- App: Spider (`spider`) — https://pipedream.com/apps/spider.md
- This page (HTML): https://pipedream.com/apps/spider/actions/scrape-new-page
- Hints: open-world
- Source: https://github.com/PipedreamHQ/pipedream/blob/master/components/spider/actions/scrape-new-page/scrape-new-page.mjs

## Description

Initiates a new page scrape (crawl). [See the documentation](https://spider.cloud/docs/api#crawl-website)

## Props

| Prop | Type | Required | Description |
|---|---|---|---|
| `url` | `string` | Yes | The URI resource to crawl, e.g. https://spider.cloud. This can be a comma split list for multiple urls. |
| `limit` | `integer` | No | The maximum amount of pages allowed to crawl per website. Default is 0, which crawls all pages. |
| `storeData` | `boolean` | No | Decide whether to store data. Default is false. |

## Run it

**MCP**

```ts
import { Client } from "@modelcontextprotocol/sdk/client/index.js"
import { StreamableHTTPClientTransport } from "@modelcontextprotocol/sdk/client/streamableHttp.js"
import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const accessToken = await pd.rawAccessToken

const transport = new StreamableHTTPClientTransport(
  new URL("https://remote.mcp.pipedream.net/v3"),
  {
    requestInit: {
      headers: {
        Authorization: `Bearer ${accessToken}`,
        "x-pd-project-id": process.env.PIPEDREAM_PROJECT_ID!,
        "x-pd-environment": "production",
        "x-pd-external-user-id": "{external_user_id}", // any stable ID for this user in your system
        "x-pd-app-slug": "spider",
      },
    },
  },
)

const mcp = new Client({ name: "my-agent", version: "1.0.0" })
await mcp.connect(transport)

const { tools } = await mcp.listTools()

// listTools() hands your model this tool's input schema, so it can
// fill the arguments itself:
const result = await mcp.callTool({
  name: "spider-scrape-new-page",
  arguments: {
    url: "URL",
    limit: 10,
  },
})
```

**TypeScript**

```ts
import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "spider-scrape-new-page",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    spider: { authProvisionId: "apn_xxxxxxx" },
    url: "URL",
    limit: 10,
  },
})

console.log(result)
```

**cURL**

```bash
curl -X POST https://api.pipedream.com/v1/connect/{project_id}/actions/run \
  -H "Content-Type: application/json" \
  -H "X-PD-Environment: production" \
  -H "Authorization: Bearer {access_token}" \
  -d '{
    "external_user_id": "{external_user_id}",
    "id": "spider-scrape-new-page",
    "configured_props": {
      "spider": { "authProvisionId": "apn_xxxxxxx" },
      "url": "URL",
      "limit": 10
    }
  }'
```

---

- App: https://pipedream.com/apps/spider.md · All apps: https://pipedream.com/apps
