why wouldn't they benchmark the accuracy against jev too?
They say they are only benchmarking public models in the blog.
Also, section 2.3: https://typesafe.ai/legal/mca
They say they are only benchmarking public models in the blog.
Also, section 2.3: https://typesafe.ai/legal/mca