LogoClawIndex
CasesSkillsAbout
LogoClawIndex

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.

Skills with capability: Extract image features

Browse skills that share this capability.

  • blip-2-vision-language - BLIP-2 vision-language task framework
    MultimodalVision-LanguageImage CaptioningVQA

    ★ 85 · Updated 2026-10-04

    Supports image captioning, visual question answering, image-text matching, feature extraction, and multimodal understanding with frozen vision and language models.

    ⚙ Generate image captions⚙ Answer visual questions⚙ Match images and text