What building a GPU cloud product in 2019–2020 taught me about AI infrastructure economics
Inference cost-per-token is a tariff problem. I first met it as V100/vGPU configurations, utilization curves and support tickets — and the lessons transfer almost one to one.
~1,200 words · AI infrastructure