Yes that is number of reasoning tokens used.
Performance increases both with larger model (Luna vs Sol)
And with more reasoning (low vs xhigh)