The Local LLM Index / In-Browser / #42
Nehanth/swarmllm
by Nehanth · In-Browser · updated 2d ago
Every device brings a slice. Together they run the whole model. Peer-to-peer LLM inference across browser tabs: a from-scratch WebGPU engine and a WebRTC runtime that split a 27B model over the devices in a room.
70
momentum
257
stars
40
forks
#42
rank
browserdistributed-inferenceinferencellmlocal-llmpeer-to-peerspeculative-decodingwebgpuwebrtcwgsl
View on GitHub →