Go China, screw America*
*within the scope of open models only
I like my Apache 2.0 licensed Gemma, and NVIDIA’s Nemotrons are decent bases for finetuning or continued pretraining, esp thanks to good documentation and tooling.
Oh, and Mira’s thinking machines lab dropped Inkling, a ~1T open weight model too.
This isn’t US vs China. This is open vs closed.
I like my Apache 2.0 licensed Gemma, and NVIDIA’s Nemotrons are decent bases for finetuning or continued pretraining, esp thanks to good documentation and tooling.
Oh, and Mira’s thinking machines lab dropped Inkling, a ~1T open weight model too.
This isn’t US vs China. This is open vs closed.