GLM-5.3-Flash lists at a fraction of its competitors’ prices, but the discount you actually get depends entirely on one undocumented default: unconfigured requests run at maximum reasoning effort, cutting the real-world price advantage over Qwen3.8-Max from 12x to 2.3x. We spent $14.75 and 2,700 API calls finding the parameter that gets it back — and the accuracy cost of pushing further.