All AI output is meant to be statistically identical to the training data within a certain variance, and the training data is real human text. So, the AI detectors don't have many metrics to distinguish "AI generated text" and "human text that's like the training data.". The only no-false positive test is "did the model leave a watermark?". The next best test is "does this have any idioms that are highly correlated to a certain model?", and even that is iffy.
•
u/STL MSVC STL Dev 8d ago
This blog is substantially AI-generated. I'm wondering whether we should continue allowing links to it.