No word on model hallucinations in the blog post.
https://artificialanalysis.ai/models/gpt-6-astra?omniscience...
The benchmark you linked to shows GPT-6 Astra having the lowest hallucination rate of all tested models.
Does hallucination matter for this application? We've moved beyond raw recall being that important, it seems like for law specifically all relevant facts will be cited and checked easily by humans.
Not like common law doesnt already have a lot of hallucination going on.