Their entire analysis is based on the assumption that Pangram 3.3 is a reliable tool for detecting AI-generated content. What if not?