★ 602 · Updated 2026-06-15
Upload, transform, and deliver images and videos with Cloudinary platform
Browse skills that use this input.
★ 602 · Updated 2026-06-15
Upload, transform, and deliver images and videos with Cloudinary platform
★ 2 · Updated 2026-05-28
Routes requests to scene skills or direct Meitu CLI tool execution
★ 34 · Updated 2026-05-28
Extract frames from videos/images, render Remotion projects, and perform specification checks and visual verification.
★ 748 · Updated 2026-05-09
Multi-speaker dialogue audio creation with Dia TTS
★ 0 · Updated 2026-03-24
Compress and optimize images, videos, audio, PDFs, and archives on macOS and Linux
★ 0 · Updated 2026-03-22
Automate TikTok content operations via Composio toolkit: upload videos, post photos, manage content, view profiles and stats
★ 816 · Updated 2026-03-19
Process audio files in multiple formats
★ 18 · Updated 2026-03-09
Enables Claude to manage TikTok content, engagement, and analytics through browser automation
★ 0 · Updated 2026-02-23
Provides video player components and media preparation tools for video streaming applications
★ 211 · Updated 2026-02-23
Transcribe audio to text using Whisper models via inference.sh CLI
★ 0 · Updated 2026-02-18
Queue job management patterns, processors, and async workflows for video/image processing
★ 602 · Updated 2026-02-13
Command line interface for accessing Z.AI capabilities including vision analysis, web search, page reading, GitHub repository exploration, and MCP tool integration
★ 84 · Updated 2026-02-13
Provides YouTube channel management and video operations through AgenCo's secure cloud gateway
★ 1 · Updated 2026-02-11
Process multimedia files for conversion, encoding, streaming, filtering and format optimization
★ 0 · Updated 2026-02-11
Processes multimodal content using Google Gemini AI
★ 602 · Updated 2026-02-11
Process multimodal content using Google Gemini API for analysis and information extraction
★ 3 · Updated 2026-02-11
Process multimedia files including video, audio conversion, streaming, filtering, and encoding
★ 1,018 · Updated 2026-02-09
Integrates with Google Gemini API for text generation, multimodal processing, function calling, and streaming capabilities
★ 748 · Updated 2026-02-09
Build multi-step content creation pipelines combining image, video, audio, and text generation