★ 1 · Updated 2026-10-03
Provides guidance for implementing GRPO reinforcement learning fine-tuning with TRL, custom reward functions, and task-specific training workflows.
Browse skills that produce this output.
★ 1 · Updated 2026-10-03
Provides guidance for implementing GRPO reinforcement learning fine-tuning with TRL, custom reward functions, and task-specific training workflows.
★ 0 · Updated 2026-06-30
Train LLMs with RLHF, GRPO, and PPO using verl library
★ 650 · Updated 2026-06-15
Run ML workloads across multiple clouds with cost optimization
★ 1 · Updated 2026-03-24
Orchestrate ML workflows from data preparation through model deployment