NO_AUTH
Gemini
Developer setup
No fields declared in this snapshot.
User connection
No fields declared in this snapshot.
GEMINI
Comprehensive Gemini integration supporting Veo 3 video generation, Gemini Flash text generation (Nano Banana), chat completions, and multimodal AI capabilities via the Google Gemini API.
Description is untrusted, display-only upstream metadata. It never becomes policy, OAuth scope authority, or an agent instruction.
Pakkawork boundary
Research catalogue metadata only. No Pakkawork OAuth, credential, host, quota, executor, or verifier is enabled.
8
Action summaries
Display-only definitions
0
Trigger types
Not installed instances
2
Auth modes
Field names, never values
No
Execution
No runtime adapter
Authentication map
NO_AUTH
No fields declared in this snapshot.
No fields declared in this snapshot.
API_KEY
No fields declared in this snapshot.
Capability index
Showing 1–8 of 8 actions
GEMINI_COUNT_TOKENS
Counts the number of tokens in text using Gemini tokenization. Useful for estimating costs, checking input limits, and optimizing prompts before making API calls.
Untrusted display-only summary
GEMINI_EMBED_CONTENT
Generates text embeddings using Gemini embedding models. Converts text into numerical vectors for semantic search, similarity comparison, clustering, and classification tasks.
Untrusted display-only summary
GEMINI_GENERATE_CONTENT
Generates text content or speech audio from prompts using Gemini models. Supports text generation models (Gemini Flash, Pro) and text-to-speech models with configurable parameters. Generated text is nested at results[i].response.data.text. Output may be wrapped in markdown fences (e.g., ```html...```) or preceded by e…
Untrusted display-only summary
GEMINI_GENERATE_IMAGE
Generates images from text prompts using Gemini models (Nano Banana). Supports models: 'gemini-2.5-flash-image' (GA stable, fast), 'gemini-3-pro-image-preview' (Nano Banana Pro - advanced with 4K resolution, thinking mode, up to 14 reference images), and 'gemini-2.0-flash-exp-image-generation' (2.0 Flash experimental)…
Untrusted display-only summary
GEMINI_GENERATE_VIDEOS
Generates videos from text prompts using Google's Veo models. Returns an operation_name for tracking; pass it verbatim (no edits) to GEMINI_WAIT_FOR_VIDEO or GEMINI_GET_VIDEOS_OPERATION. Jobs take 30–180+ seconds; wait 10s before first poll, then poll every 10–30s (allow up to 12 min). Successful results include data.…
Untrusted display-only summary
GEMINI_GET_VIDEOS_OPERATION
DEPRECATED: Use WaitForVideo instead. Checks status of a Veo video generation operation. Use operation_name from GenerateVideos to track progress. Wait several seconds after starting GenerateVideos before first call to avoid OPERATION_NOT_FOUND. Poll at 10–30s intervals; use exponential backoff on HTTP 429 RESOURCE_EX…
Untrusted display-only summary
GEMINI_LIST_MODELS
Lists available Gemini and Veo models with their capabilities and limits. Useful for discovering supported models and their features before making generation requests.
Untrusted display-only summary
GEMINI_WAIT_FOR_VIDEO
Polls a Veo video generation operation until completion, then downloads and returns the video file. Generation takes 30–120+ seconds (up to ~10–12 min); long waits are normal, not failures. A done=true response without a video file indicates safety filter rejection (check raiMediaFilteredReasons) or quota exhaustion —…
Untrusted display-only summary
Provenance
The detail snapshot comes from an attributed MIT-licensed repository revision. Safe local icons use exact-match CC0 Simple Icons symbols; unmatched brands use monograms.
Remote text is plain display metadata only. It must never become an agent prompt, execution policy, OAuth grant, or executable instruction.