GLM was twice as verbose running the Artificial Analysis benchmark. So it ends up being more expensive

Not really. Gemini 3.6 Flash actually cost $0.01 more per task, compared to GLM 5.2.

https://artificialanalysis.ai/models/gemini-3-6-flash

But to run the entire benchmark it cost $727 with Gemini 3.6 Flash and $925 with GLM-5.2, $198 (21.4%) less. I tend to look at the cost to run the whole index rather than the weighted average cost per task.