LogoClawIndex
CasesSkillsAbout
LogoClawIndex

google-gemini-api - Google Gemini API Integration for Multimodal AI Generation

Integrates with Google Gemini API for text generation, multimodal processing, function calling, and streaming capabilities

Tags

Updated: 2026-02-09

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • generate text content
  • process images
  • handle video files
  • process audio inputs
  • parse PDF documents
  • execute function calls
  • stream responses
  • configure thinking mode
  • manage chat conversations

Inputs

  • Google API key
  • text prompts
  • image files
  • video files
  • audio files
  • PDF documents
  • function schemas
  • system instructions
  • model parameters

Outputs

  • generated text responses
  • multimodal responses
  • streamed content chunks
  • function call results
  • usage metadata

Requirements

  • Google Gemini API key
  • Node.js runtime (for SDK)
  • @google/genai SDK v1.27+
  • JavaScript/TypeScript environment
  • Internet connectivity

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
AI
API integration
multimodal
text generation
Google Cloud
generate text content
process images
handle video files
process audio inputs
Google API key
text prompts
image files
generated text responses
multimodal responses
streamed content chunks