OP's point here is that the overall approach of restricting output token space and using parallel prompts to produce concurrent results and taking the most relevant ones isn't something novel to Jev (not saying there's nothing novel, but a facsimile can be created at the application layer using any small, fast model)
I still don't get the point of jev....it's basically an optimized models/runner on really short context and output?
I get the point, and it's nice, but I think the "Jev" naming is confusing (and it could be legally dangerous).
What’s novel is how fast and cheap Jev is while maintaining quality. If they’re trying to say they made the same thing, that is likely incorrect. Getting the same result 100x faster is in fact a breakthrough technology.