You can partially tell by the tokeniser; which gives you some hint into the training corpus mix.
</div> is four Gemma4 tokens, but one Qwen3.6 token.
Where do you find this information for each model?
Where do you find this information for each model?