I'd be surprised if it were anything other than just an LLM only equivalent of Gemma some 4B param model, they can't be spending much money on it because each query essentially needs to be cheaper than the ad revenue per search.
Could they even afford 4B dense parameters touching every token? That seems like a LOT of compute compared to generating the classic Google SERP.
Could they even afford 4B dense parameters touching every token? That seems like a LOT of compute compared to generating the classic Google SERP.