Run LocalMode Bench on your device

Pick a suite, keep the tab visible, and press Run. Models download once (they cache for next time), the protocol runs warmups and timed iterations per the published methodology, and when the run completes the result publishes automatically to the public leaderboard (switch publishing off to keep a run local; you can always export the raw JSON).

Configure the run
SmolLM2 135M (MLC q0f16)webllm · 78 MB · q0f16
unavailable
SmolLM2 135M (GGUF Q4_K_M)wllama · 70 MB · Q4_K_M
unavailable
BGE Small EN (ONNX, WebGPU)transformers-webgpu · - · q8
unavailable
BGE Small EN (ONNX, WASM)transformers-wasm · - · q8
unavailable
BGE Small EN (GGUF Q8_0)wllama · 35 MB · Q8_0
unavailable
Universal Sentence Encoder (MediaPipe)mediapipe · -
unavailable

Probing device capabilities…

0 lanes · est. download - (cached models skip the download)

Models download only when you press Run. Keep this tab visible and your device plugged in - hidden tabs invalidate timed runs. With publishing on, the result uploads to the open dataset automatically when the run completes: timings, device environment, and the generated text for the fixed public prompts. No personal data. Turn the toggle off to keep the run local (JSON export only).