Three-LLM is a WebGPU-powered inference engine built on Three.js. It allows for running large language models directly in the browser using hardware acceleration.
HOW THIS AFFECTS YOU
●
builderYou can deploy low-latency, client-side LLM inference without relying on backend GPUs.