Compute

Community submitted benchmarks for Local AI

526 community benchmarks across 37 models and 16 chips, from 9 contributors. Decode and prefill tok/s by model, quantisation and chip, measured with BaseRT or llama.cpp.

Visit Compute

More in Dev Tools

Opening Liz…