I think GLM-5 is a regression on interactive Python coding though, been using it almost daily and GLM 4.7 before that. The most likely culprit is DSA — and I conclude it's not straightforward to apply. Likely V4 manages better, but there will be tradeoffs.