No API key, no server call at inference time.
SmolLM2-135M-Instruct (180MB) runs entirely in this browser tab via WebGPU.
Load it once, then try airplane mode (without reloading the page) โ
extraction still works.
โก WebGPU accelerated๐ Works with wifi off๐ Zero network calls at inference