LogoClawIndex
CasesSkillsAbout
LogoClawIndex

agentic-eval - Agentic Evaluation Patterns

Patterns and techniques for evaluating and improving AI agent outputs.

Tags

Updated: 2026-09-16

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • Implement self-critique and reflection loops
  • Build evaluator-optimizer pipelines
  • Create test-driven code refinement workflows
  • Design rubric-based evaluation systems
  • Design LLM-as-judge evaluation systems
  • Measure and improve response quality

Inputs

  • Task descriptions
  • Evaluation criteria
  • Code specifications
  • Evaluation rubrics

Outputs

  • Refined agent outputs
  • Structured JSON evaluation scores
  • Fixed code implementations

Requirements

    Source

    • Spec: SKILL.md

    ClawIndex

    OpenClaw Skills & Use Case Index

    ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

    Index

    Skills·
    Cases

    Meta

    About·
    Disclaimer·
    Email·
    GitHub
    © 2026 ClawIndex All Rights Reserved.
    evaluation
    reflection
    evaluator-optimizer
    llm-as-judge
    quality-control
    Implement self-critique and reflection loops
    Build evaluator-optimizer pipelines
    Create test-driven code refinement workflows
    Design rubric-based evaluation systems
    Task descriptions
    Evaluation criteria
    Code specifications
    Refined agent outputs
    Structured JSON evaluation scores
    Fixed code implementations