← Back to the wire

Running an LLM in Your Browser: Verifying WebGPU, Model Hashes, and Local Inference Without a Server

AnnouncementProductAug 30, 2026

WebGPU reached stable cross-browser support in 2025, enabling LLMs from Hugging Face's Transformers.js v4 to run fully in-browser. Open-weights models (Llama, Phi, Qwen) in the 1B–8B range with 4-bit quantization occupy 0.8–5 GB, fitting consumer hardware. CapyToolkit provides free browser tools for verification: a hash verifier for SHA-256 checksums, a token counter for tokenizer cross-checks, and a fingerprint inspector. Users should also monitor DevTools' Network tab for outbound leaks.

Receipt № 16561 source · awaiting confirmation ◐

Evidence

1source· awaiting independent confirmation

No score is assigned. Sources and their independence are shown in the citation chain below.

Citation chain · 1 source

Hugging FaceCompanyQwenModelLlamaModelCapyToolkitCompanyPhiModel
Canonical: https://capytoolkit.com/blog/developer-tools/running-llm-browser-verifying-webgpu-model-hashes-local-inference/