logoalt Hacker News

tarrudatoday at 12:13 PM1 replyview on HN

Hopefully it will be open weights and have the same architecture and size as the current v4 flash vision, which is probably the best LLM that can be run on 128G devices.


Replies

fluoridationtoday at 12:26 PM

Interesting, I had assumed it'd be too large to fit. What quant and context size are you running?

show 1 reply