Testing inference at all system scales.
Independent benchmarks of system inference performance, covering hardware & inference software across all scales: datacenter nodes/racks, workstations, and mobile devices.
Choose an inference scale
We benchmark systems of all sizes on the throughput, latency, cost and power use metrics that matter for decision-making. We customize our inference benchmarks for each of the three size tiers below.
Datacenter
Compare AI accelerator performance at node, rack and cluster scale.
Workstation
Coming late 2026
Compare consumer- and prosumer-tier system performance for running open models at your desk.
Portable devices
Coming Aug 2026
Compare portable consumer and developer device performance running small open models.