The Local LLM Index / Inference Engines / #80
SergiioB/intel-arc-pro-b70-inference-cookbook
by SergiioB · Inference Engines · updated today
Open recipes, engine patches, and benchmark harnesses for LLM inference on Intel Arc Pro B60/B70 (Battlemage, Xe2). MoE 35B at 160 t/s decode / 7.5K t/s prefill single-stream, 27B at 50~ t/s decode / 1.7K t/s prefill single stream. vLLM XPU MTP unlocked. Muse Glimmer recipe added!!
65
momentum
126
stars
9
forks
#80
rank
battlemageintel-arcintel-arc-b70intel-gpullama-cppllm-inferencelocal-aimoemtpspeculative-decodingsyclvllm
View on GitHub →