you can pretty much run frontier-level models locally now with GLM 5.3 Flash and Qwen 3.8 Flash Next so this pricing feels very geared towards squeezing people who are still dependent on cloud subscription models only and don't know how to divert grunt work to cheaper models automatically with the 'smarter' frontier models acting as orchestrators and refinement
https://this.os.isfine.org/blog/posts/chinese-flash-models-a...