Testing inference at all system scales.

Independent benchmarks of system inference performance, covering hardware & inference software across all scales: datacenter nodes/racks, workstations, and mobile devices.

Choose an inference scale

We benchmark systems of all sizes on the throughput, latency, cost and power use metrics that matter for decision-making. We customize our inference benchmarks for each of the three size tiers below.

Datacenter

Compare AI accelerator performance at node, rack and cluster scale.

Workstation

Coming late 2026

Compare consumer- and prosumer-tier system performance for running open models at your desk.

Portable devices

Coming Aug 2026

Compare portable consumer and developer device performance running small open models.