Breaks down why inference pricing looks nothing like compute cost, from batching and KV-cache economics to who actually makes money serving models.