LogoClawIndex
CasesSkillsAbout
LogoClawIndex

prompt-engineer-toolkit - LLM Prompt Engineering and Testing Toolkit

A/B test prompts, version them, and ensure regression safety for production LLM workflows

Tags

Updated: 2026-03-23

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • Run A/B prompt tests
  • Score prompt outputs
  • Track prompt versions
  • Compare prompt versions
  • Create prompt templates
  • Run regression tests
  • Audit prompt quality
  • Calculate prompt metrics

Inputs

  • Test cases JSON file
  • Baseline prompt file
  • Candidate prompt file
  • LLM runner command
  • Prompt name
  • Change notes

Outputs

  • A/B test results
  • Prompt version history
  • Version diff report
  • Quality metrics report
  • Prompt audit report

Requirements

  • Python 3 runtime
  • Local script execution
  • LLM CLI access
  • JSON configuration support
  • Local filesystem access

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
prompt engineering
LLM optimization
A/B testing
prompt versioning
marketing
AI workflows
Run A/B prompt tests
Score prompt outputs
Track prompt versions
Compare prompt versions
Test cases JSON file
Baseline prompt file
Candidate prompt file
A/B test results
Prompt version history
Version diff report