Production LLM bills rarely match the input-token rate on a pricing page. Output tokens, reasoning tokens, non-prod traffic, verbosity, and long-context surcharges compound invisibly. This piece maps the five multipliers and shows what we measured on DigitalOcean Serverless Inference.