★ 3 · Updated 2026-09-29
Train and optimize sparse Mixture of Experts models with DeepSpeed or HuggingFace, including routing, load balancing, expert parallelism, and inference optimization.
Browse skills that produce this output.
★ 3 · Updated 2026-09-29
Train and optimize sparse Mixture of Experts models with DeepSpeed or HuggingFace, including routing, load balancing, expert parallelism, and inference optimization.
★ 650 · Updated 2026-09-24
Provides 9 reinforcement learning algorithms to create, train, and manage learning plugins for autonomous agents in AgentDB.
★ 2 · Updated 2026-09-23
Provides guidance for AI system design, MLOps architecture, scalable ML infrastructure, and AI platform engineering.
★ 0 · Updated 2026-06-30
Train LLMs with RLHF, GRPO, and PPO using verl library
★ 1 · Updated 2026-06-30
Flexible RL training library for large language models