View as Markdown
Google Gemini icon

Google Gemini ACTION

Generate Content from Text and Image

Generates content from both text and image input using the Gemini API. See the documentation
  • Action
  • Writes data
  • API key
  • SDK
  • MCP

IMPLEMENTATION

Call this tool

Connect a user's Google Gemini account once, then configure and run Generate Content from Text and Image from your backend or agent.

import { PipedreamClient } from "@pipedream/sdk"

const pd = new PipedreamClient({
  projectId: process.env.PIPEDREAM_PROJECT_ID!,
  clientId: process.env.PIPEDREAM_CLIENT_ID!,
  clientSecret: process.env.PIPEDREAM_CLIENT_SECRET!,
  projectEnvironment: "production",
})

const result = await pd.actions.run({
  id: "google_gemini-generate-content-from-text-and-image",
  externalUserId: "{external_user_id}", // any stable ID for this user in your system
  configuredProps: {
    google_gemini: { authProvisionId: "apn_xxxxxxx" },
    model: "Model",
    text: "Prompt Text",
  },
})

console.log(result)

SCHEMA

Inputs

Pipedream supplies the connected account. Your application provides the operation-specific values below. Dynamic inputs are resolved against that user's account.

Generate Content from Text and Image inputs
Property Type Description
model Model string
The model to use for content generation
Required Dynamic
text Prompt Text string
The text to use as the prompt for content generation
Required
responseFormat JSON Output boolean
Enable to receive responses in structured JSON format instead of plain text. Useful for automated processing, data extraction, or when you need to parse the response programmatically. You can optionally define a specific schema for the response structure.
Optional
history Conversation History string[]
Previous messages in the conversation. Each item must be a valid JSON string with text and role (either user or model). Example: {"text": "Hello", "role": "user"}
Optional
safetySettings Safety Settings string[]
Configure content filtering for different harm categories. Each item must be a valid JSON string with category (one of: HARASSMENT, HATE_SPEECH, SEXUALLY_EXPLICIT, DANGEROUS, CIVIC) and threshold (one of: BLOCK_NONE, BLOCK_ONLY_HIGH, BLOCK_MEDIUM_AND_ABOVE, BLOCK_LOW_AND_ABOVE). Example: {"category": "HARASSMENT", "threshold": "BLOCK_MEDIUM_AND_ABOVE"}
Optional
mediaFiles Media File Paths or URLs string[]
A list of file paths from the /tmp directory or URLs for the media to process.
Required
syncDir SyncDir dir
Optional

REFERENCE

Tool details

Behavior hints are published with the component in the Pipedream registry and surface as MCP tool annotations, so an agent can reason about a tool before it calls it.

Registry key
google_gemini-generate-content-from-text-and-image
Version
1.0.3
App
Google Gemini
Authentication
API key
Read-only
No
Destructive
No
Open world
Yes