LogoClawIndex
CasesSkillsAbout
LogoClawIndex

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.

Skills tagged: LLM evaluation

Browse skills that share this tag.

  • llm-as-a-judge - Build LLM Evaluators for Automated Quality Assessment
    LLM evaluationquality assessmentautomated evaluationTPR

    ★ 9 · Updated 2026-05-28

    Build and deploy LLM evaluators for automated Pass/Fail quality assessment of LLM pipeline outputs

    ⚙ write judge prompt⚙ split labeled data⚙ measure TPR/TNR
  • promptfoo - LLM Evaluation Framework for Testing and Comparing Outputs
    LLM evaluationtesting frameworkprompt engineeringquality assurance

    ★ 602 · Updated 2026-03-09

    CLI tool for testing and comparing LLM outputs using evaluation configurations and assertions

    ⚙ run evaluation configurations⚙ create test cases⚙ debug evaluation runs