LogoClawIndex
CasesSkillsAbout
LogoClawIndex

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.

Skills tagged: Quantization

Browse skills that share this tag.

  • serving-llms-vllm - High-throughput LLM serving framework
    vLLMInference ServingPagedAttentionContinuous Batching

    ★ 0 · Updated 2026-09-22

    High-throughput LLM serving engine supporting OpenAI compatible API, quantization, and tensor parallelism.

    ⚙ Deploy production LLM APIs⚙ Run offline batch inference⚙ Serve quantized LLM models
  • gguf-quantization - GGUF Quantization for Efficient Model Inference
    GGUFQuantizationllama.cppCPU Inference

    ★ 0 · Updated 2026-06-30

    Convert and quantize models to GGUF format for efficient CPU/GPU inference

    ⚙ convert model to GGUF⚙ quantize GGUF model⚙ run model inference
  • llama-cpp - Local GGUF Inference & Model Discovery
    llama.cppGGUFQuantizationHugging Face Hub

    ★ 0 · Updated 2026-05-11

    Run local GGUF models and discover Hugging Face repositories for llama.cpp

    ⚙ run local models⚙ search Hugging Face Hub⚙ discover GGUF files
  • llama-cpp - Local GGUF Inference & HF Discovery
    llama.cppGGUFQuantizationHugging Face Hub

    ★ 17 · Updated 2026-05-11

    Local GGUF inference, quant selection, and Hugging Face Hub model discovery for llama.cpp

    ⚙ Run local models⚙ Search Hugging Face Hub⚙ Find GGUF files
  • llama-cpp - Local GGUF Model Inference
    GGUFLocal InferenceHugging FaceQuantization

    ★ 1 · Updated 2026-05-11

    Run GGUF models locally and discover models from Hugging Face Hub

    ⚙ discover Hugging Face repos⚙ download GGUF files⚙ run GGUF models locally