The Local LLM Index / In-Browser / #42

Nehanth/swarmllm

by Nehanth · In-Browser · updated 2d ago

Every device brings a slice. Together they run the whole model. Peer-to-peer LLM inference across browser tabs: a from-scratch WebGPU engine and a WebRTC runtime that split a 27B model over the devices in a room.

70
momentum
257
stars
40
forks
#42
rank
browserdistributed-inferenceinferencellmlocal-llmpeer-to-peerspeculative-decodingwebgpuwebrtcwgsl
View on GitHub →