The Local LLM Index / Quantization & Formats / #200
SqueezeAILab/SqueezeLLM
by SqueezeAILab · Quantization & Formats · updated 2y ago
[ICML 2024] SqueezeLLM: Dense-and-Sparse Quantization
32
momentum
724
stars
52
forks
#200
rank
efficient-inferencelarge-language-modelsllamallmlocalllmmodel-compressionnatural-language-processingpost-training-quantizationquantizationsmall-modelstext-generationtransformer
View on GitHub →