What metric should a data center operator use to compare AI infrastructure efficiency across different GPU generations when the hardware costs and power draw are both changing?