View as Markdown
Speak AI icon

Speak AI ACTION

Upload Media

Upload an audio or video file to Speak AI for transcription and analysis, from a publicly reachable URL or an AWS signed URL. Processing is asynchronous, so use the New Automated Transcription (Instant) trigger to act on the result. See the documentation.
  • Action
  • Writes data
  • OAuth
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's Speak AI account once, then configure and run Upload Media from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "speak_ai-upload-media",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    speak_ai: { authProvisionId: "apn_xxxxxxx" },
    name: "Name",
    url: "URL",
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Upload Media inputs
Property Type Description
name Name string
Name of the media file
Required
url URL string
Public URL or AWS signed URL
Required
mediaType Media Type string
Type of media file (audio or video)
Required
folderId Folder ID string
A Speak AI folder ID, e.g. 905c208f1c07. The folder to upload to, or to retrieve files from. Returned as folderId by List Folder ID Options
Required Dynamic
description Description string
Description of the media file
Optional
tags Tags string[]
Optional metadata tags for the media file upload
Optional

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
speak_ai-upload-media
Version
0.0.4
App
Speak AI
Authentication
OAuth
Read-only
No
Destructive
No
Open world
Yes