Local LLM Index
/ local
← All tools
◐
The Local LLM Index
/ Quantization & Formats / #226
leafspark/AutoGGUF
by leafspark · Quantization & Formats · updated 8mo ago
automatically quant GGUF models
27
momentum
226
stars
19
forks
#226
rank
View on GitHub →
More in Quantization & Formats
Quantization & Formats
#1
unsloth
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek
Quantization & Formats
#4
llamafile
Distribute and run LLMs with a single file.
Quantization & Formats
#13
qwen38-27b-rtx3090
Qwen3.8-27B on a single RTX 3090 with vLLM: ~1,000 tok/s at 64 concurrent (int8 tensor-cor
Quantization & Formats
#18
Soup
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.