logoalt Hacker News

simianwordstoday at 9:54 AM0 repliesview on HN

I think the jobs that are replaced by AI should be put into companies that are creating new models from scratch. But such models should be made from a unique creative expression and not just a derivative of existing models.

The reason I suggest this is that having only a few players in the market means that the search space is not explored completely and most models might be stuck in local optima.

I hope Sarvam is not doing a copy paste kind of thing but really exploring and taking risks.

But question is: how are they getting the training data? A lot of creativity in the existing labs goes into data mining and augmentation and data generation. Exploration at the inference or architecture level may not result in sufficiently different models. The world doesn’t need another Qwen