★ 92 · Updated 2026-09-09
Runs local GGUF inference, selects quantizations, and discovers llama.cpp-compatible models and files on Hugging Face.
Browse skills that share this tag.
★ 92 · Updated 2026-09-09
Runs local GGUF inference, selects quantizations, and discovers llama.cpp-compatible models and files on Hugging Face.
★ 0 · Updated 2026-06-30
Convert and quantize models to GGUF format for efficient CPU/GPU inference
★ 0 · Updated 2026-05-11
Run local GGUF models and discover Hugging Face repositories for llama.cpp
★ 17 · Updated 2026-05-11
Local GGUF inference, quant selection, and Hugging Face Hub model discovery for llama.cpp
★ 1 · Updated 2026-05-11
Run GGUF models locally and discover models from Hugging Face Hub
★ 602 · Updated 2026-03-26
Fine-tune LLMs with Unsloth using 4-bit quantization and LoRA
★ 18 · Updated 2026-03-10
Secondary local LLM inference engine for running GGUF models directly, loading LoRA adapters, benchmarking, and custom model serving