logoalt Hacker News

ProllyInfamousyesterday at 5:56 PM1 replyview on HN

Exactly; when I first got my RTX 5070 Ti (16gb, to game with!!!, upgrading from VEGA56), I loaded then-latest Qwen3.6 (~30B, cannot remember exactly). My only prior LLM experience was with models <8gb, primarily llama3.1.

My technical-expert twin played around with these LLMs, for about an hour, and then correctly reasoned "it's able to be WRONG, faster."

This seems apt. My next LLM machine will be closer to 96gb+ vRAM.


Replies

selectodudeyesterday at 6:35 PM

Once I get some kind of settlement after getting beaten up by a cop my first purchase will be some RTX Pro 6000s.