★ 0 · Updated 2026-09-17
Converts photos of handwritten letters into installable TTF and web font files.
Browse skills that share this tag.
★ 0 · Updated 2026-09-17
Converts photos of handwritten letters into installable TTF and web font files.
★ 166 · Updated 2026-09-15
Routes image, document, and OCR tasks to specialized execution tools and returns standard JSON output for text models.
★ 0 · Updated 2026-09-15
Generate text, images, video, speech, and music, analyze images, and perform web search via MiniMax AI using the mmx terminal CLI.
★ 0 · Updated 2026-09-11
Configures custom or OpenAI-compatible vision providers such as Kimi, DeepSeek, or Azure in Hermes Agent.
★ 2,284 · Updated 2026-09-10
Converts local images, image URLs, or clipboard images into text descriptions by running the bundled vision.js script.
★ 0 · Updated 2026-05-28
Take screenshots, detect UI elements, click, scroll, type, and press keys to control macOS applications through their graphical interface.
★ 59 · Updated 2026-05-28
Control macOS GUI applications through visual interaction using screenshots, clicks, scrolling, and typing
★ 2 · Updated 2026-05-28
Integrate AI capabilities into Stacks applications with multiple provider drivers
★ 1 · Updated 2026-05-28
AI/LLM integration with multiple providers, RAG, image generation, and personalization features
★ 620 · Updated 2026-05-28
Integrate AI capabilities into Stacks apps with multiple drivers, image generation, RAG, embeddings, and personalization features
★ 0 · Updated 2026-03-22
Captures photos and videos from USB webcam, sends via WhatsApp, optionally describes content with vision model
★ 816 · Updated 2026-03-10
A lightweight tool that directly calls MiniMax Coding Plan APIs for web search and image understanding using pure JavaScript.
★ 0 · Updated 2026-02-11
Processes multimodal content using Google Gemini AI
★ 602 · Updated 2026-02-11
Process multimodal content using Google Gemini API for analysis and information extraction