★ 0 · Updated 2026-09-22
High-throughput LLM serving engine supporting OpenAI compatible API, quantization, and tensor parallelism.
Browse skills that produce this output.
★ 0 · Updated 2026-09-22
High-throughput LLM serving engine supporting OpenAI compatible API, quantization, and tensor parallelism.
★ 57 · Updated 2026-06-30
Deploys Prometheus, Grafana, and Node Exporter monitoring stack using Docker Compose
★ 650 · Updated 2026-03-20
Set up comprehensive observability for Windsurf integrations with metrics, traces, and alerts