Even if hardware capacity increases, it seems clear it uses way more tokens, so I don't expect parity with other competitors on that front.
On the other hand I expect K3 future refinements to be massive and more efficient.
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.
The speed at which tokens are crunched, even on the same hardware, differs between models as well. Using more tokens is only a problem if they are processed at the same speed as with a comparison model.