WattThe / Data center / PUE / efficiency
PUE tells you how much of the power bill actually reaches the servers. The AI-era companions — TFLOPS per watt and tokens per watt — tell you what the compute produces per unit of energy.
Power Usage Effectiveness is total facility power divided by IT power. A PUE of 1.35 means that for every 100 kW reaching servers, another 35 kW is spent on cooling, power conversion losses, and building systems. The theoretical floor is 1.0; the global fleet averages around 1.5–1.6, while hyperscale builds with economization and liquid cooling report 1.1–1.2. The metric's blind spot is that it says nothing about whether the IT power itself does useful work — an idle cluster at PUE 1.1 is still waste.
That's why AI operators increasingly track compute productivity per unit of energy. TFLOPS per watt measures raw arithmetic capability against the full facility draw, useful for comparing hardware generations — each new accelerator family has roughly doubled it. Tokens per watt (or per watt-hour) is the inference-era metric: how much model output the whole facility produces per unit of energy, which folds in model efficiency, batching, quantization, and utilization, not just silicon.
Together the three numbers form a chain: PUE tells you how efficiently power reaches compute, TFLOPS/W tells you how capable that compute is, and tokens/W tells you how much value comes out. Anyone underwriting an AI facility should be asking for all three.
The industry fleet average sits around 1.5–1.6. Modern purpose-built facilities achieve 1.2–1.4, and the best hyperscale sites with free cooling and liquid loops report 1.1 or lower. Anything above 2.0 signals a legacy facility with major efficiency headroom.
No — by definition total facility power includes IT power, so PUE is always at least 1.0. Claims below 1.0 usually involve on-site generation or heat-reuse accounting that steps outside the standard definition.
It measures how many model tokens (units of AI output) a system produces per watt of power, folding hardware efficiency, model optimization, and utilization into one productivity number. It is becoming a standard way to compare inference deployments.
Estimates are for planning and education, not engineering design or financial advice. Verify rates with your utility and confirm electrical work with a licensed engineer or electrician.
Join the WattThe list for state rate updates, new calculators, and a deeper walkthrough of your numbers. Commercial or data center project? Reply to any email for a feasibility review.
No spam. One or two emails a month, unsubscribe anytime.