★ 22 · Updated 2026-10-05
Deploy and serve Kimi-K2.6 with INT4 QAT or NVFP4 using vLLM or SGLang in Docker on eight RTX PRO 6000 Blackwell GPUs.
Browse skills that share this tag.
★ 22 · Updated 2026-10-05
Deploy and serve Kimi-K2.6 with INT4 QAT or NVFP4 using vLLM or SGLang in Docker on eight RTX PRO 6000 Blackwell GPUs.
★ 0 · Updated 2026-10-03
Quantizes diffusion DiTs with NVIDIA ModelOpt, converts FP8 or NVFP4 exports for SGLang Diffusion, and validates quality and performance.
★ 1 · Updated 2026-09-24
Provides guidance and workflows for LLM post-training with RL using the slime framework integrating Megatron-LM and SGLang.
★ 650 · Updated 2026-05-11
Enterprise-grade RL training framework for large MoE models with FP8/INT4 support