LogoClawIndex
CasesSkillsAbout
LogoClawIndex

evaluate-presets - Systematically test Ralph hat collection presets using scripts.

Evaluates Ralph hat collection presets by running test scripts, logging session metrics, and verifying hat routing performance.

Tags

Updated: 2026-09-18

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • Evaluate single preset
  • Evaluate all presets
  • Extract session metrics
  • Validate hat routing
  • Generate summary report

Inputs

  • Preset name
  • Backend selection
  • Test task definitions

Outputs

  • Evaluation log files
  • Metrics JSON files
  • Markdown summary report
  • Execution status codes

Requirements

  • Cargo build system
  • yq command-line tool
  • Bash execution environment

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
testing
evaluation
presets
ralph
benchmarking
Evaluate single preset
Evaluate all presets
Extract session metrics
Validate hat routing
Preset name
Backend selection
Test task definitions
Evaluation log files
Metrics JSON files
Markdown summary report