Model is great. Have to try with oh my py --prewalk with either Kimi as planner and GLM 5.3 as implementer or GLM 5.3 as planner and Flash as implementer.
They have a new shiny data center with Chinese GPUs, so hopefully they will be able to handle the demand. Their subscriptions are meh, in particular the Flash model has not that much usage. The other thing I wanted to try is Groq Qwen 3.8-27B at 450 tps to see if it's able to get work done faster.
I feel it's very important to play outside of the walled gardens, I feel they will pull the rug very soon, and I need to get job done while waiting hardware prices to go down...