k3 costs will go down at least 3x within a week of the weights dropping.
we'll get new quants, dspark speculators, distills and optimized kernels
as long as there are near frontier models available there will be inference providers selling them at or below cost of inference in attempt to get market share.
I have not seen that kind of significant drop with GLM 5.2 yet? so curious why you think it will happen for K3.
This is a very large model. Much larger (3x) than GLM. The resources to run it are very expensive.