Yeah, a good rule of thumb is that the weights take up ~100% of the size of the model, so 100B bytes (8-bit quant) would be, well, 100GB and a 4-bit quant would be half that.
Oh, that is a useful rule to know! Thanks!
Oh, that is a useful rule to know! Thanks!