★ 650 · Updated 2026-06-30
Provides guidance for training LLMs with RL using verl library
Browse skills that use this input.
★ 650 · Updated 2026-06-30
Provides guidance for training LLMs with RL using verl library
★ 0 · Updated 2026-06-30
Train LLMs at scale using verl with PPO, GRPO, and other RL algorithms
★ 1 · Updated 2026-06-30
Flexible RL training library for large language models supporting multiple algorithms and backends
★ 0 · Updated 2026-06-30
Train large language models with reinforcement learning algorithms using the verl library
★ 66 · Updated 2026-06-30
Pipeline for embedding skills into local models via LoRA fine-tuning on Apple Silicon
★ 650 · Updated 2026-03-26
Train LLMs to communicate natively using the Slipstream protocol
★ 650 · Updated 2026-02-14
Develops and fine-tunes AI models for specific tasks and domains.