Good observation! It would have to be offset with O(140k) queries to the model, which is, well, unlikely.
If it's about the latency / flow disruption, spending a few hours once could easily be worth it if the result is actually good enough to skip googling/retries.
Just like with OSS in general, being able to distribute it is what makes the effort worthwhile.
This particular example is maybe a niche, but 1400 people can use a few hundred queries in a reasonable amount of time.