LogoClawIndex
CasesSkillsAbout
LogoClawIndex

verl-rl-training - verl: LLM RL Training Framework

Flexible RL training library for large language models

Tags

Updated: 2026-06-30

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • train models
  • run rollout
  • compute reward
  • configure algorithms
  • swap backends
  • execute multi-turn rollout
  • train vision-language models
  • enable LoRA

Inputs

  • Base model
  • Training dataset
  • Reward function
  • Training config

Outputs

  • Trained model checkpoints
  • Training metrics

Requirements

  • Python 3.8+
  • verl>=0.3.0
  • torch>=2.0.0
  • ray>=2.41.0
  • vllm>=0.8.2
  • transformers>=4.40.0
  • GPU cluster
  • HuggingFace model

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
Reinforcement Learning
RLHF
GRPO
PPO
Post-Training
Distributed Training
Megatron-LM
vLLM
train models
run rollout
compute reward
configure algorithms
Base model
Training dataset
Reward function
Trained model checkpoints
Training metrics