★ 602 · Updated 2026-06-30
Provides guidance for training LLMs with RL using verl library
Browse skills that use this input.
★ 602 · Updated 2026-06-30
Provides guidance for training LLMs with RL using verl library
★ 0 · Updated 2026-06-30
Train LLMs at scale using verl with PPO, GRPO, and other RL algorithms
★ 1 · Updated 2026-06-30
Flexible RL training library for large language models supporting multiple algorithms and backends
★ 0 · Updated 2026-06-30
Train large language models with reinforcement learning algorithms using the verl library
★ 11 · Updated 2026-06-30
Deep analysis of techniques for compressing large models into smaller efficient ones via distillation
★ 3,789 · Updated 2026-06-15
Catalog common PyTorch mistakes and provide battle-tested training patterns and optimization techniques
★ 3 · Updated 2026-05-28
High-level PyTorch framework with automatic distributed training and minimal boilerplate
★ 602 · Updated 2026-03-26
Train LLMs to communicate natively using the Slipstream protocol
★ 602 · Updated 2026-03-26
Fine-tune LLMs with Unsloth using 4-bit quantization and LoRA
★ 602 · Updated 2026-03-22
Detect systematic bias in AI systems and ensure predictions do not vary unfairly across protected attributes
★ 15 · Updated 2026-03-21
Complete machine learning environment with PyTorch for model training and GPU acceleration
★ 2 · Updated 2026-03-10
Provides guidance for scaling ML training across multiple GPUs and nodes with parallelism strategies