LogoClawIndex
CasesSkillsAbout
LogoClawIndex

bench-open-model - Benchmark Open-Weight Models on AWS GPUs

Sizes, provisions, serves, benchmarks, reports on, and tears down AWS GPU infrastructure for vLLM-servable Hugging Face text and multimodal models.

Tags

Updated: 2026-10-02
AWSGPU benchmarkingvLLMHugging Faceopen-weight modelsthroughput testinglatency testingmultimodal inference

Capabilities

Estimate model VRAM needsRank AWS GPU instancesProvision EC2 GPU infrastructureServe models with vLLM

Typical Inputs

Hugging Face model IDAWS credentialsAWS account and region

Typical Outputs

VRAM sizing resultsGPU instance recommendationsThroughput and latency results

What this skill does

  • Estimate model VRAM needs
  • Rank AWS GPU instances
  • Provision EC2 GPU infrastructure
  • Serve models with vLLM
  • Run cache-honest load sweeps
  • Measure throughput and latency
  • Write benchmark reports
  • Teardown benchmark resources

Inputs

  • Hugging Face model ID
  • AWS credentials
  • AWS account and region
  • Benchmark workload shape
  • Explicit launch approval
  • HF access token
  • GPU instance preference
  • vLLM image version

Outputs

  • VRAM sizing results
  • GPU instance recommendations
  • Throughput and latency results
  • Benchmark report file
  • Benchmark state file
  • vLLM serving endpoint
  • Provisioned AWS resources
  • Deleted CloudFormation stack

Requirements

  • Python 3
  • AWS CLI
  • AWS launch permissions
  • SSH and SCP access
  • vLLM architecture support
  • Dedicated non-production AWS account
  • Available AWS GPU capacity

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.