logoalt Hacker News

arbugetoday at 4:52 PM1 replyview on HN

I think what they're really getting hooked on is the lowest cost provider.

Which makes the Muse 1.3 launch this week particularly interesting, although to get the low cost version you do need to agree to share data with Meta.


Replies

AnotherGoodNametoday at 5:19 PM

I actually think they’re hooked on models they can fine tune.

You can’t further train the closed models. The open models can be fine tuned for your company. Big companies fine tune models on all the internal systems and documentation, not just through .md files (you’d blow up the context trying it that way) but actual fine tuning of open weights models. A low tier but open weights model actually beats frontier models when you do this for a specific task.

I think the frontier providers need to have a way to isolate instances (bedrock style?) and allow fine tuning to compete. Big companies are absolutely fine tuning models right now and getting better results than even the best frontier models for their use cases.