logoalt Hacker News

dotancohen • yesterday at 5:49 PM • 1 reply • view on HN

It would be great if we could correct Jev's incorrect answers, even on a separate endpoint. Let me tell it what Jev got wrong.

What type of head is that? What type of model is that head part of?


Replies

tgluck • yesterday at 7:11 PM

Not today, but Interesting idea. The main motivation was a drop-in for an existing Jev setup, so the only teacher right now is Jev and the audit measures agreement with Jev. A correction would have to become a second label source that overrides Jev's for that input.

The head is a multinomial logistic regression: one linear layer plus softmax on top of a frozen sentence-embedding model (bge-small by default, swappable). That head is the entire local model, the encoder is off the shelf and never changes.

➕ show 1 reply