Browser-local inference · No API keys

Local AI Lab

A collection of experiments exploring local AI inference in the browser using open-source models from Hugging Face.

Experiments7
RuntimeWebGPU / WASM
BackendNone
Data sent to a serverNever

How it runs

The purpose of this repository is to document what I learn while experimenting with AI models that can run locally, without relying on external inference APIs, API keys, or sending data to a server. Each experiment is both a short technical write-up and a working demo.

Each page focuses on a specific AI experiment. The first interaction shows the model's download size, then downloads quantized ONNX weights from the Hugging Face CDN and caches them in the browser. Inference happens in a Web Worker: WebGPU when the browser can initialize it, with WASM as a CPU fallback.

Pick an experiment, load it once, and try the samples. The Python and JavaScript snippets under each demo show the same pipeline outside this site.

Experiments

Contact

github · ignaciojsoler@gmail.com