Charlie Ruan leads WebLLM, which runs open models locally in your browser
Charlie F. Ruan leads WebLLM, an open-source runtime that uses WebGPU, WebAssembly, and compiler-generated artifacts to run language models locally in a browser without an inference server. WebLLM, puβ¦