The Local LLM Index / Inference Engines / #66
flashrt-project/FlashRT
by flashrt-project · Inference Engines · updated today
FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
65
momentum
460
stars
58
forks
#66
rank
cudacuda-kernelsgr00tgr00t-n1-6-3bjetsonjetson-orinjetson-thormotuspipi05qwenqwen3-6
View on GitHub →