logoalt Hacker News

polyomino • today at 10:27 PM • 1 reply • view on HN

Even though this is way more expensive than backprop, could a hybrid approach where you fine tune an existing checkpoint that's been backpropped unlock further gains? It would be cool to apply this to different stages and see if that affects the learning trajectory


Replies