> they don't value correctness
I spend ~10-20 tokens on correctness and codebase cleanliness for every 1 I spend on new features.
How do you measure that? And how do you know it's effective, or that it's even working, if you're not reviewing the results?
How do you measure that? And how do you know it's effective, or that it's even working, if you're not reviewing the results?