The Local LLM Index / Quantization & Formats / #146

john-rocky/apple-silicon-llm-bench

by john-rocky · Quantization & Formats · updated 3d ago

Reproducible on-device LLM benchmarks for Apple Silicon (iPhone 17 Pro, M4 Max): Apple Core AI, MLX, llama.cpp, LiteRT-LM and Core ML on the same model and harness, every number with its quantization and capture session; hybrid Mamba-2 models (Nemotron-3 Nano, Granite-4.0-H, Falcon-H1) included.

52
momentum
67
stars
5
forks
#146
rank
apple-core-aiapple-siliconbenchmarkcore-aicoreaicoremlgranitehybrid-mambaiosios-27iphone-17-prollama-cpp
View on GitHub →