logoalt Hacker News

Auracletoday at 4:58 PM1 replyview on HN

Sure, but shouldn’t the programs to run the LLMs go “the user has this much vram and the model is this size, so I’ll start with sensible defaults based on that”?

You could override, obviously.


Replies

zargontoday at 5:05 PM

Yes, llama.cpp does that.