> More capable than Mythos 5 in some areas, less capable in others; overall slightly more capable.
This sounds like it might be a Mythos finetune for some specific task.
EDIT: After reading some more reading, it looks like model 2 might be an AI research fine tune based off the section 3.4.3 CoBench