★ 0 · Updated 2026-05-11
Adapt and debug Hugging Face or local models for vLLM on Ascend NPU
Browse skills that use this input.
★ 0 · Updated 2026-05-11
Adapt and debug Hugging Face or local models for vLLM on Ascend NPU
★ 0 · Updated 2026-05-11
Adapt and debug models to run vLLM on Ascend NPU
★ 60 · Updated 2026-05-11
Adapt and debug models for vLLM on Ascend NPU.
★ 650 · Updated 2026-05-11
Generates deployment artifacts for ML models with APIs, containerization, monitoring, and A/B testing infrastructure
★ 22 · Updated 2026-03-23
Manage GPU VRAM sharing across services with OOM retry and auto-unload