I find these explanations that drag in things like "dice" incomprehensible. What dice?
What the article is trying to say:
"For AI-generated text, watermarking usually works by secretly biasing which words/tokens the model prefers while it writes."
Simple. Understandable. No "dice" or "leaning" or other convoluted explanations
Is it actually biasing specific words over the course of the text? Or something less direct in terms of how any given next token is predicted? Seems like biasing words would be too big of a tell.
“Provenance” seems to be the flavor of the day.