Our inference cost is growing faster than revenue and we have already optimized the model tier so what are operators using at the infrastructure level to actually reduce cost per token?