The Local LLM Index / Quantization & Formats / #123

jjang-ai/jangq

by jjang-ai · Quantization & Formats · updated 8d ago

JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon

57
momentum
225
stars
28
forks
#123
rank
apple-siliconggufjang-quantizationllamacppllmmlxmlxllmomlxomlx-alternativequantization
View on GitHub →