when you're waiting for network or LLM inference the raw performance doesn't matter at all
You would think raw performance wouldn't be a problem given most of whats happening is waiting for network calls and streaming tokens.
But modern bloat manages perfectly well to make apps that wait for network calls run poorly enough to give you a bad experience.
This is so very, very incorrect.
yeah fair enough, my entire point is not about the application itself but the contradiction on using superlative terms for all points but a compromising/normal term for one. Like if performance is not revelant why include it in the list of benefits.
This line of thinking I feel like assumes it's the only program running on your computer. Using less of my CPU and memory means my computer can do more things in parallel, or even run more instances of the harness. My laptop is sweating when I got 5+ claude code sessions running.