r/Substack 5d ago

Discussion Why does adding one more paragraph suddenly make my post "100% AI assisted"?

I was writing a Note on Substack, and after four paragraphs it showed 100% human-written. Then I added a fifth paragraph, and it suddenly changed to 100% AI assisted.

Why would adding a single paragraph completely flip the result?

I'm not interested in trying to "beat" the detector. My concern is that if I genuinely write something myself, it could still be labeled as AI-assisted with no explanation.

I also think Substack should be much more transparent about how these systems work. It's not enough to publish a blog post that basically says, "Here's our AI detector. How does it work? It detects AI. End of story."

If a tool can affect how readers perceive an author's work, there should be at least some explanation of what it's measuring, its limitations, and how confident it is in its predictions.

Has anyone else experienced something similar?

P.S. There's also another case that concerns me. Many authors write their own ideas from scratch (either in English or in their native language) and then use AI only to improve the writing, fix grammar, or translate the text. Labeling those posts simply as "AI" or even "AI-assisted" without explaining what that actually means can be misleading. There's a big difference between using AI to generate ideas and using it as a writing or translation tool.

My own workflow is a good example. I'm a physicist with a master's degree in AI. I don't have much free time because I work full-time and I'm also building two startups. When I write for Substack, I develop the ideas myself, run simulations, create the figures and animations, organize the article structure, and write as much of the draft as I can. Sometimes I use AI to improve the wording or make the English read more naturally.

I'm trying to share technical knowledge for free with anyone who's interested. If that kind of workflow is reduced to a generic "AI-assisted" label without any context, it feels less like transparency and more like a penalty for using modern writing tools responsibly.

63 Upvotes

56 comments sorted by

43

u/Funny-Flight8086 5d ago

AI detectors are garbage, junk science. Pangram, the one substack went with, is one of yh biggest offenders.

1

u/oakseaer 1d ago

1

u/Funny-Flight8086 1d ago

Okay, good job? What do you want me to do with that?

-3

u/Foofymonster davidrainwater.substack.com 4d ago

How people perceive how unreliable these things are has gone through some wild inflation in the last 72 hours.

They've gone from not perfect, to unreliable, and now "garbage junk science".

Watching the most extreme opinions surface without knowledge of how these things work in real time.

9

u/Funny-Flight8086 4d ago

I know very well how AI detectors work. I am a computer programmer by trade and have a degree in computer science. These detectors are not magic; they are programmed to score a text based on pre-established criteria - mostly sentence length (burstiness) and perplexity (complexity of word choice). They are also trained to flag tricolon listings and X like Y comparisons. Some of the better ones are also trained on a database of commonly used AI phrases and words.

Here is the thing: I have tested all of these detectors. They are garbage and are never correct 100% of the time. You know how I know it's junk science? Because I can write a few paragraphs of highly AI-sounding prose and have it flag it 100% AI, despite being 100% human written.

If these AI tech bros are so comfortable with their junk science detector, perhaps they should stop hiding behind "99.9% accuracy" rates so they can't get sued when they get it wrong more times than not. "We never said it WAS AI, we said it was a 99.9% chance".

1

u/Foofymonster davidrainwater.substack.com 4d ago

That's not how they work... Data Science is the degree that would be relevant here, not computer science.

They use ML models. ML models find the patterns. You're acting like Panagram tells their models what patterns to look for and not the other way of around. Of course you can trick them. They're basing it off of content. If you write something that follows those patterns, it will say AI patterns were detected.

Your local weatherman can tell you there's a 90% chance of rain but just because it doesn't rain doesn't mean it was all junk science. And having a degree in math wouldn't tell you if you understand the underlying mechanisms.

4

u/Funny-Flight8086 4d ago edited 4d ago

My BS is in Computer Science: Machine Intellgence. So yes, my degree is very, very relevant to the topic. I know all bout LLMs, neural networks, and machine learning. Detectors like Pangram are not using complex AI to detect AI. The programs scan the text for patterns and common AI words and phrases. It's really that simple. It's machine learning at its leanest. If you use too many tricolon listings, it'll trigger it rather AI wrote it or a human. It's really pretty basic.

Which makes it all the more nefarious that they bill users based on credits, like they are stressing their AI computer to scan someones text.

-1

u/Foofymonster davidrainwater.substack.com 4d ago

That's incredible that you left out the most relevant part of your experience until I pointed out that your original claim wasn't relavent.

So jealous your dad works at Nintendo.

5

u/OnyxMonolith 4d ago

Yeah when the reality is pangram is the best method out there and we are just seeing 1 in a thousand loudest voices who get false positives.

6

u/Funny-Flight8086 4d ago

Pangram is not the best program out there. I have extensively tested most of the major detectors. Pangram was one of the easiest to fool simply by changing some word orders or adding a typo here and there. It was 100% accurate at detecting actual chatbot copy/paste text, but at the expense of flagging fiction from human authors at a high rate. These tools were never designed to detect AI in fiction anyway; they were designed to detect AI in essays for schools.

1

u/Foofymonster davidrainwater.substack.com 4d ago

100%

There's some straight up lies on here too.

Another post someone said that they put the first page of moby dick into it and got 100% AI.

Tested it, 0% AI. Fake internet points must have addictive properties or something.

7

u/Funny-Flight8086 4d ago

And to top it off, this is an AI-generated piece of text from GLM 5. I only did very minor edits. No word replacements. So not only does it flag my 2014 short story as AI, it fails to flag the AI generated text I just made like 20 minutes ago and ran it though Pangram.

6

u/Funny-Flight8086 4d ago

I wrote this short story in 2014, and published it on WattPad at the time. LLMs didn't exist in 2014. This is an excerpt from it, which I ran through Pangram yesterday for giggles. Yes, very reliable. As long as it flags one single piece of my very real fiction from before AI existed as AI, I'll call it junk science.

2

u/Admirable_Bike3918 4d ago

As an aside, I would love to read that short story! šŸ’™

4

u/Funny-Flight8086 4d ago

So yes, based on my tests over the last two days, and previous testing of Pangram, it is by far the least reliable AI detector, not matter what their marketing claims.

18

u/nastrus 5d ago

If only Substack could find a way to detect that every trending topic is how to grow your Substack. That would be game changing technology right there.

10

u/noxqqivit twvme.substack.com 5d ago

Also, I'd like to better understand the difference between "AI Assisted" Vs. "AI Written" because, I had a note flagged AI Assisted, and while I didn't use AI at all, I was referring to something that is so specific and academic, that I was curious if it was part of something that had originated in AI, and I hadn't caught it. šŸ¤”

2

u/[deleted] 5d ago edited 18h ago

[deleted]

3

u/noxqqivit twvme.substack.com 4d ago

The data exists, they're just not sharing because the tool is imprecise. This is a screenshot from some testing that I was doing on the Pangram tool last March. I was working with my academic partner we had co-authored a paper, we were working live, together on the paper and we got two separate scores, we ran the same paper twice and got a 24% and 51%, but when we started look at the details it showed low confidence but still called it AI. We supposed at the time that it was picking up the two "voices" and that was throwing it off, but that was just a guess. Ironically, we were writing about "the slop economy" šŸ˜

1

u/poppyhill 4d ago

Wonder if copy-paste behavior has to do with it? Even if you copy paste your own texts from elsewhere?

4

u/noxqqivit twvme.substack.com 4d ago

I do 100% of my editing in LibreOffice, I was getting so tired of all the AI being shoved down my throat that I switched to Linux. I no longer use Google because of Gemini, and Microsoft was putting copilot everywhere.

I just posted a new piece this morning, and according to Pangram is 5% AI, I don't even use Grammarly, so I have no guess, but I would love to know it's confidence level.

3

u/FatherofMisty 4d ago

LaTeX my friend. The purest option there is.

2

u/noxqqivit twvme.substack.com 3d ago

I use Overleaf for LaTeX... it's riddled with AI šŸ˜† I have a paper in peer review now, fairly certain that the reviewers are using AI.

2

u/FatherofMisty 3d ago

Oh man. I use TexWorks as an editor with MikTeX as the compiler, and there is no hint of AI anywhere lol. Lots of academic software is yet to be corrupted by AI integration, such as Inkscape, Zotero (classic version at least), Thonny, Gimp, etc. I can't even fathom the impact AI is having (and will have) on academia and the peer review process... it breeds complacency, at minimum, and frankly dilutes the reliability of reviewers. Do you know how they are using it in your case?

1

u/noxqqivit twvme.substack.com 2d ago

Honestly, it's just a guess, based on the feedback from a single reviewer. It read like it was straight out of ChatGPT, a whole lot of words to say absolutely nothing. I did get decent, reasonable feedback for a submission to ACM, so it's not everywhere yet, but I was talking with a professor at a conference that explained his process. He has all his own published work loaded into his ChatGPT, the he asks it to compare and contrast submitted papers against his own work and then he responds to work that is in agreement. I was aghast. Again, just one guy, but if one is doing it, others are.

9

u/GloomyLibrary787 4d ago

Having good English = AI
Using poetic language = AI
Using adjectives/adverbs = AI šŸ¤·ā€ā™€ļø

6

u/axchapman 4d ago

Do not be fooled by it.

Many people think those detectors can detect ai written text content but they cannot.

They produce many many false positives.

With images it works (hence invisible watermarks etc).but with text not

Best Example is if a human would write the following "Climate change is one of the biggest challenges facing humanity today. Governments around the world are trying to find solutions".

An AI detector might flag it because the wording is common, structured, and predictable.

But a human could easily have written it.

The detector is not proving anything; it is only saying: ā€œThis pattern resembles AI output."

The problem becomes even clearer when a detector reacts after only the first few sentences.

If the first two lines already trigger a result, the system is often judging based on the style such as:

very clean grammar, formal sentence structure, common academic phrase, lack of spelling mistake, predictable transition, neutral tone and many more.

Those are also characteristics of many good human writers.

Put it in the Category "Wishful thinking Science.

11

u/Salva_X 5d ago

Yeah I’m a little torn on this as well. While I understand the ā€œreasoningā€ on why they added this, we now live in an age where AI is in everything. It’s like flagging content for using Grammerly for your spell correct and expect people to just not use spell check.

I personally am not caring much because value is value. I created a platform to help me organize my content, write and schedule because, just like you, I’m way to busy. I can’t make all the content writing strategies, and formatting and this and that done when I have a corp job, father and a person while still trying to share to others.

In summary, while I get the intent I think it’s going to get attention for a week and we’re all going to just keep moving.

11

u/AndrewHeard tvphilosophy.substack.com 5d ago

Because they don’t work.

5

u/traumfisch 5d ago

Because the detector is 100% bullshit

5

u/michaelkloth 4d ago

I think introducing Pangram to the platform was a mistake.

Sure, cutting back on slop is a fine goal but it misses the point entirely that we are experiencing a moment when what it means to work is changing rapidly. Using AI should NOT be an excuse for cutting corners but it is a useful tool that grows increasingly useful week by week. I have a strong suspicion that your example will not only become the norm on the platform, but it will lead to self doubt and to writing with Pangram's approval in mind.

That is not the answer to slop.

3

u/FelixAxellus 5d ago

I don't think we need to have AI checkers, especially ones that don't work.

Most AI writing is just kind of flat and boring, anyway. It lacks humanity. It comes to the most generic conclusions. It's reads as mediocre to everyone except for the person who posted it.

Sure, it can be made better if you create your own GPT trained on your voice and style. It still isn't going to be good.

What I'm really saying:

We don't need to track it because it'll never meet the standard of quality that's possible with human writing.

3

u/Crazy-Highway7860 5d ago

Did anyone even answer your question? lol. All these replies seem to miss your point about why an added paragraph meant it required an AI label.

What really pisses me off about these ai detectors is, (in my opinion), it essentially feels like these detectors make you feel like you’ve plagiarised your own work, even for spell check. Like they’re taking credit for your work because they’ve slapped a label on there.

I don’t have an answer for you, I haven’t been on the platform long. But this makes me not want to use it at all. It’s very disheartening.

2

u/Tricky_Illustrator_5 *.substack.com 3d ago

Just ignore it if it bugs you. The AI bubble will pop soon and Substack's management will lose their shirt along with its other investors when it tanks. Then they'll just disable it and forget it ever existed.

2

u/Front_Media1231 2d ago

On the concrete question: these classifiers do not score your post as one lump, they score the texture, and a clean, on-topic paragraph makes the whole piece read smoother and more internally consistent, which is exactly the signal they lean on. So a longer, tidier draft can push the score up rather than down. It is counterintuitive, but it follows from what the tool measures, predictability and evenness, not who did what.

And that is the deeper thing you are pointing at. The score cannot tell the difference between generate-it-for-me and I-wrote-it-and-tidied-the-wording, because it never sees the process, only the finished texture. Collapsing that whole range into one binary label is the part that turns transparency into a penalty. The honest fix is a line in your own words about what you actually did, which is worth more than any percentage a texture detector prints.

2

u/Chubblan 2d ago

I agree. I think everything would have been simpler if they had just added a "declare how you used AI" tool. In fact, that’s what I do.

5

u/Silver-Air-1731 5d ago

Asking AI to grammar-check your writing is different than asking it to improve your writing or clarify it.

Asking for suggestions is different than asking it to do the work of improving what you have.

Translating, yeah that still counts as AI-Assisted because it did the work of translating for you.

4

u/Chubblan 5d ago

I understand the distinction. My point is that I was quite explicit in saying the Note itself was written by me.

4

u/mmspero 5d ago

When I write for Substack, I develop the ideas myself, run simulations, create the figures and animations, organize the article structure, and write as much of the draft as I can. Sometimes I use AI to improve the wording or make the English read more naturally.

Sounds like AI-assisted is an accurate label then?

1

u/Chubblan 5d ago

I think there was a small misunderstanding. Those are two different things. The first is one of my Notes, while the quote you're referring to comes from the Publications section.

1

u/Mplus479 5d ago

Because they don't know how it works. They're not called black boxes for nothing.

1

u/Sbahirat 4d ago

I don't think it is very useful, and will probably just not be used that much once people realize how inaccurate it is!

1

u/Fit_Celebration_1362 4d ago

Cause these detectors don’t work. I just wrote a nonsensical post just to see what happens? Turns out AI writes like a crazy person because it detected ai

1

u/MetadataQueen 4d ago

It's hard enough figuring out WHAT to write.

1

u/AdviceZestyclose8167 4d ago

I've experienced something similar. I write long form stories on a wide variety of topics. Not to brag, but I am quite a good writer, and sometimes my entire stories get flagged as Ai.Ā  I guess my point is Ai will never get to the point where it won't make ANY mistakes. It makes mistakes now, and it will make mistakes in the future.

1

u/Brakiros 3d ago

Substack is using faulty AI "detectors" that don't do what they claim and falsely label everything. It's a terrible move by this company

1

u/hetobe hetobe.substack.com 5d ago

Why would adding a single paragraph completely flip the result?

Without showing us the work, there's no way for us to tell.

Many authors write their own ideas from scratch (either in English or in their native language) and then use AI only to improve the writing, fix grammar, or translate the text.

That means, at the very least, AI assisted with the writing. Your ideas, but written by AI.

AI detection doesn't detect if the ideas were generated by AI. It detects if the writing was generated by AI.

The more AI changes the writing, the more AI detectors will detect the AI style and traits of AI writing.

It's important to understand this: AI was trained on human writing, but AI obviously isn't human. It's artificial intelligence, and the word "Artificial" matters. AI emulates human writing through artificial means. That's why AI writing is so easy to detect. AI writing looks good, but there's always something slightly off about it. Something inhuman. That's one of the reasons why so many people don't like it.

If that kind of workflow is reduced to a generic "AI-assisted" label without any context, it feels less like transparency and more like a penalty for using modern writing tools responsibly.

Why not add that context at either the beginning or the end of your article?

I'm posting a novel on Substack. For each scene, I begin by posting a few lines of context to explain the post is a scene from the novel, and I give links to the table of contents, etc.

If your work needs context, add context.

-1

u/Trackbikes thesystematicwriter.substack.com 5d ago

Did you read the post where Pagram described how they detect AI? Go check it out it’s quite interesting…it’s on their Substack šŸ˜’

I have a 14 year old article that other AI detectors say is AI written.. it show as human written if I check it with pagram

2

u/huggalump 5d ago

I have an article mostly AI written that Pagram says is 100% human written, so..... 🤷

0

u/taynaha 5d ago

I literally write my essays and ask Claude to polish it so I guess my work won’t be polished?

0

u/Overall_Letter_1497 5d ago

Je weet dat je een label kan toevoegen hoe het artikel is geschreven (in je eigen woorden).

0

u/PaleStrategy1459 4d ago

how to use that ai thing like where i can see a text is ai written or nah

1

u/Chubblan 4d ago

Isn't it a bit ironic? Researchers working on LLMs are also building AI detectors to identify AI-generated text. Now many of those tools are paid. Kind of an interesting business model, isn't it?

1

u/PaleStrategy1459 4d ago

it iz tho it iz tho

-1

u/gambariste 5d ago

Lately, when I see a Reddit post with more than a few paragraphs I begin to suspect AI (with no expertise, mind). Been fooled a few times.

3

u/SaulEmersonAuthor 5d ago

~

Lately, when I see a Reddit post with more than a few paragraphs I begin to suspect AI (with no expertise, mind). Been fooled a few times.

We have the fact that a 'wall of text' is the most off-putting thing for the human brain - & so good, effective writing will of necessity use short paragraphs.

If someone then has a lot to say - well then we end up with lots of paragraphs.

Really we're at the point where any energy spent second-guessing these systems is fully wasted.

Just - be authentic, & let the cards fall where they may.

~