The Local LLM Index / Quantization & Formats / #146
john-rocky/apple-silicon-llm-bench
by john-rocky · Quantization & Formats · updated 3d ago
Reproducible on-device LLM benchmarks for Apple Silicon (iPhone 17 Pro, M4 Max): Apple Core AI, MLX, llama.cpp, LiteRT-LM and Core ML on the same model and harness, every number with its quantization and capture session; hybrid Mamba-2 models (Nemotron-3 Nano, Granite-4.0-H, Falcon-H1) included.
52
momentum
67
stars
5
forks
#146
rank
apple-core-aiapple-siliconbenchmarkcore-aicoreaicoremlgranitehybrid-mambaiosios-27iphone-17-prollama-cpp
View on GitHub →