All inference happens locally in your browser via ONNX Runtime Web (WASM) — no server, no GPU, no upload.
Description excerpted from the original listing, which is linked below.