LogoClawIndex
CasesSkillsAbout
LogoClawIndex

llama-cpp - Local GGUF Inference & Model Discovery

Run local GGUF models and discover Hugging Face repositories for llama.cpp

Tags

Updated: 2026-05-11

Capabilities

Typical Inputs

Typical Outputs

What this skill does

  • run local models
  • search Hugging Face Hub
  • discover GGUF files
  • generate text
  • create chat completions
  • generate embeddings
  • start llama-server
  • run llama-cli
  • list available GGUFs
  • query tree API
  • read repo files

Inputs

  • GGUF model file
  • Hugging Face repository
  • quantization variant
  • context length
  • GPU layers count
  • chat messages
  • text prompt
  • repository URL

Outputs

  • generated text
  • chat response
  • text embedding
  • server endpoint
  • GGUF file list
  • CLI command
  • discovery result

Requirements

  • llama.cpp
  • llama-cpp-python
  • CPU or GPU
  • RAM or VRAM
  • Linux/macOS/Windows

Source

  • Spec: SKILL.md

ClawIndex

OpenClaw Skills & Use Case Index

ClawIndex is an ecosystem-driven index of OpenClaw skills and real-world use cases.

Index

Skills·
Cases

Meta

About·
Disclaimer·
Email·
GitHub
© 2026 ClawIndex All Rights Reserved.
llama.cpp
GGUF
Quantization
Hugging Face Hub
CPU Inference
Apple Silicon
GPU Support
Edge Deployment
Model Discovery
run local models
search Hugging Face Hub
discover GGUF files
generate text
GGUF model file
Hugging Face repository
quantization variant
generated text
chat response
text embedding