logoalt Hacker News

ChickeNEStoday at 3:11 AM0 repliesview on HN

> And show me an api provider that allows me to run 10x agents concurrently for 5 days straights .

Any of them on a Max/Pro plan as long as you are smart about model selection? That's my main objection to local inference, I'd need a whole rack of GPUs to do as many things in parallel that I can do for $400 a month. I do plan on setting up some local inference hardware, but...RAM and GPU prices alone are $$$$