WebLLM: high-performance in-browser LLM inference engine
Enables developers to build privacy-preserving, client-side AI applications with hardware acceleration and familiar APIs.
AI Summary
WebLLM is a high-performance in-browser LLM engine that runs models locally using WebGPU with full OpenAI API compatibility.
