★ 248,179 · Updated 2026-03-19
Optimizes LLM inference on NVIDIA GPUs with quantization and multi-GPU scaling
Browse skills that share this capability.
★ 248,179 · Updated 2026-03-19
Optimizes LLM inference on NVIDIA GPUs with quantization and multi-GPU scaling
★ 0 · Updated 2026-02-23
Deploy local or open-source AI models with model selection, quantization, Ollama/vLLM setup, and integration with ModelSelector