logoalt Hacker News

phreezayesterday at 5:10 PM1 replyview on HN

This is a pretty stunning result. The time series looks really convincing. Is the way the detector itself is trained orthogonal to this or could there be some "leakage" in that the pre-chatgpt text is in the (positive) training data?


Replies

dopamine_daddyyesterday at 5:11 PM

I tried my best to avoid leakage. If you're curious about how I trained the detector I have a writeup on it: https://unslop.run/blog/how-our-ai-text-detector-works

FYI this is all relatively new so there might be lots of issues and iterations coming.

show 1 reply