★ 92 · Updated 2026-09-09
Runs local GGUF inference, selects quantizations, and discovers llama.cpp-compatible models and files on Hugging Face.
Browse skills that share this tag.
★ 92 · Updated 2026-09-09
Runs local GGUF inference, selects quantizations, and discovers llama.cpp-compatible models and files on Hugging Face.
★ 650 · Updated 2026-06-30
Implement Metal/MPS kernels for PyTorch operators on Apple Silicon
★ 0 · Updated 2026-06-30
Convert and quantize models to GGUF format for efficient CPU/GPU inference
★ 0 · Updated 2026-06-30
Implement Metal/MPS kernels for PyTorch operators using native c10/metal infrastructure
★ 18 · Updated 2026-06-15
Implements JAX-style SplitMix64 PRNG on Apple Silicon using MLX with GPU acceleration for deterministic color generation
★ 0 · Updated 2026-05-11
Run local GGUF models and discover Hugging Face repositories for llama.cpp
★ 18 · Updated 2026-05-11
Local GGUF inference, quant selection, and Hugging Face Hub model discovery for llama.cpp
★ 1 · Updated 2026-03-25
Write PyTorch and MLX code for Apple Silicon M-series hardware without CUDA
★ 816 · Updated 2026-03-21
High-performance audio library for Apple Silicon with text-to-speech and speech-to-text