The article analyzes the costs of running AI inference at scale, specifically looking at the costs of input processing and output generation. It suggests that input processing is relatively cheap, while output generation is more expensive. The author estimates that the cost of input processing is around $0.001 per million tokens, while output generation costs around $3 per million tokens. This cost asymmetry has implications for the profitability of different AI applications.