The Local LLM Index / Quantization & Formats / #123

huaxin0/FunASR-GGML

by huaxin0 · Quantization & Formats · updated 5d ago

C++ speech recognition inference engine using GGML — CPU/CUDA GPU, real-time microphone streaming, single GGUF model file, no Python dependency

55
momentum
125
stars
16
forks
#123
rank
View on GitHub →