logoalt Hacker News

woahtoday at 7:38 PM0 repliesview on HN

Huge omission. This requires fine tuning.

> Zero-shot vs. Fine-tuning: Out-of-the-box base models score ~0.35 on the typed-decisions benchmark (near random). The 0.766 score is achieved by fine-tuning on the benchmark's train split. Treat Laya as a fast foundation model to specialize, not as an omniscient zero-shot oracle.