Their fear-mongering about GLM 5.3 got me to try it out. Its very good, I'll only go back to Claude if GLM isn't available (it forgot how to do tool calls yesterday).
Interestingly enough, it seems to compact at about 10% of the 1mn context, which makes sense if they're trying to run profitably.
I recommend trying Coralbricks with GLM due to their cheap input cache prices. GLM can get quite expensive elsewhere.