logoalt Hacker News

simonwtoday at 4:03 PM2 repliesview on HN

Fine tuning LLMs has turned out to be mostly not worth the effort, but I wonder if fine tuning Jev-style models will turn out to be a whole lot more useful.


Replies

edottoday at 4:41 PM

But why? Jev-style models seem useful for "I have no clue what my incoming distribution looks like but I need to give some sort of answer". If I know what my incoming distribution looks like I'll just upload a CSV of that into ChatGPT and ask it to fit a basic ML model on my data.

alexmolastoday at 4:15 PM

If you want calibrated probabilities you'll be forced to fine-tune it