Testing inference at all system scales.
Independent benchmarks of system inference performance, covering hardware & inference software across all scales: datacenter nodes/racks, laptops and workstations, and mobile phones.
Choose an inference scale
We benchmark systems of all sizes on the throughput, latency, cost and power use metrics that matter for decision-making. We customize our inference benchmarks for each of the three size tiers below.
Datacenter
Compare AI accelerator performance at node, rack and cluster scale.
Mobile phones
Coming Aug 2026
Compare mobile phone performance running small open models.
Laptops & workstations
Coming late 2026
Compare consumer- and prosumer-tier system performance for running open models at your desk.