Testing inference at all system scales.
Independent benchmarks of system inference performance, covering hardware & inference software across all scales: datacenter nodes/racks, workstations, and mobile devices.
Choose an inference scale
We benchmark systems of all sizes on the throughput, latency, cost and power use metrics that matter for decision-making. We customize our inference benchmarks for each of the three size tiers below.
Datacenter
Compare AI accelerator performance at node, rack and cluster scale.
Portable devices
Coming Aug 2026
Compare portable consumer and developer device performance running small open models.
Workstation
Coming late 2026
Compare consumer- and prosumer-tier system performance for running open models at your desk.