logoalt Hacker News

_doctor_lovetoday at 4:26 PM4 repliesview on HN

Easy prediction: LLMs will get shrunk down further and further until GenAI is just something that ships on a chip as part of your hardware. In the future it will seem quaint that we needed a network connection to talk to our LLM.

Adoption of open-source models to my mind is a similar step in that direction. In all cases, the goal is to become untethered from a mercurial vendor.


Replies

honrtoday at 4:54 PM

Like taalas.com (very recently acquired by AMD), or cerebras.ai (whole wafer is a chip)? As you said, I also think that is one of the main direction many companies (and academia) is moving to.

hparadiztoday at 6:45 PM

We're gonna start baking in models like TTS with thousands of voices available in any language as a chip on device. They just need to hit 99% accuracy and then it's a done deal.

wnmurphytoday at 4:58 PM

Yeah, I'm looking forward to this actually.

https://chatjimmy.ai/ blew my mind at how fast etched model weights can be.

For on-device LLMs, there's a point of diminishing returns, meaning you don't need to have the latest frontier model for most operations.

show 1 reply
andriy_kovaltoday at 6:14 PM

for every small GenAI model there will be larger model or cluster of models which are smarter than small model