The Local LLM Index / Quantization & Formats / #123
jjang-ai/jangq
by jjang-ai · Quantization & Formats · updated 8d ago
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
57
momentum
225
stars
28
forks
#123
rank
apple-siliconggufjang-quantizationllamacppllmmlxmlxllmomlxomlx-alternativequantization
View on GitHub →