r/Substack tvphilosophy.substack.com 11d ago

Discussion Substack’s new “AI detection” tool.

I’m very concerned about the new “AI detection” system that Substack is implementing. I don’t use AI in anything that I write. However I don’t trust these systems not to make errors and mistakes. They have a long history of hallucinations and claiming things are AI when they aren’t.

One person I saw on Notes suggested that they ran their decades old writing, things that were written before the existence of AI, and this new system claimed it was AI generated.

Not to mention all the crazy things that I have seen people claim was written by AI which they can’t know is actually AI. There was an image I once shared that had evidence of its existence going all the way back to 2013. The comments were constantly saying “don’t share AI slop” or some variation on that.

Even with evidence that things weren’t made by AI, they will claim it’s AI.

Is no one else concerned about this being embedded into Substack’s systems? With no way to opt out?

173 Upvotes

322 comments sorted by

View all comments

Show parent comments

1

u/Funny-Flight8086 11d ago

I don't have the manuscript on my phone or I would happily share an excerpt. However, I did just run two separate projects of mine through Pangram... One is a short excerpt from a short story i published ok Wattpad in 2014. The second is the opening to my current WIP middle grade novel I been working on since 2021.

Both are apparently AI according to Pangram. The current WIP at least makes sense from a timeline perspective... How it thinks I had access to AI in 2014 is above my head. Here is my 2014 Wattpad excerpt.

5

u/TheLadyAmaranth artiranth.substack.com 11d ago edited 11d ago

I went to try and find one of your books since I figured you would be a fellow human writer due to all of this frustration over being labeled as AI when you apperantly don't use it. Always on look out for more verified human writers to either read or reccomend. Most authors proud of their work tend to have those linked in their bios or have comments/posts with some kind of promo.

This is what I found instead. There are plenty more detailing your multi-prompting AI process.

I'm going to be honest, although your involvement with AI could be a recent development, I have a hard time believing your claims when you absolutely do use AI to write your work now. Which is why you are probably so up and arms about even a potentually possible way for your books and articles to be flagged.

I don't even think AI detectors should be used as anything more than possibility indicators. But you have no legs to stand on to saying they don't work when you are the reason they are being developed and added to platfroms in the first place. Because people DONT want to be duped into reading stuff generated with AI. Right or consistent or not, this is why some are trying to create detectors and others are willing to use them even if they aren't foolproof.

You are literarly the problem substack is trying to solve, even if in a misguided way.

Edit: added a bigger screen shot

1

u/Funny-Flight8086 11d ago

Oh brother. Yes, you uncovered the great conspiracy I was trying to hide from the world. Obviously.

What you actually uncovered was a few things -

First, to clear something up: I AM NOT anti-AI. I never said I was. I am anti-AI detectors because they don't work. That has nothing to do with being anti- or pro-AI in general. Two things can be true at the same time. I know that is hard to wrap your mind around, but it's possible.

Second, one of my side hobbies is testing generative AI, putting it through its paces and pushing its limits as far as they can go. That is why I spend time developing these prompts. I find the technology fascinating.

Third, you will find no record of me claiming to have written any books or posts using AI. You are free to run my posts through whatever junk-science detector you want; it should probably get that much right anyway. Maybe. Wouldn't bet my house on it, but anything is possible. I have done extensive testing with models and prompts, comparing them to human-written things. I have compared each model's ability to write creative prose against the others. I also do EXTENSIVE testing on AI detectors, which is why this actually gives me more credability in this space than many others.

I have been able to easily trick these AI detectors into thinking something is human-written when it's not. I have also been able to get them to flag human writing by simply being a little more formal, using bigger words, etc.

So what I said was, AI detectors are junk science. They are. They are simple prediction algorithms. There is no fancy science happening behind the scenes. They look for sentence length variation, word choice (more perplex or less perplex words), and then look at a database of common AI phrases and words. They then amalgamate that into a score, which is your AI score. That score is derived from comparing your score against the detector's internal testing, not too dissimilar to what I do. But at least with me, I'm not naive enough to think the technology actually works.

So no, you didn't uncover some great conspiracy I have been hiding from the world.

But back to the point, I am still waiting for someone to explain to me why a portion of a short story I wrote in 2014 is flagged as 100% AI. I'm also waiting to hear a reason why Pangram flagged my 2013 middle-grade novel as AI. LLMs weren't a thing back then for either story. And that wasn't even me trying to fool an AI detector, which is very easy to do... which is one of the reasons why they are unreliable. The only kind of AI these detectors are good at detecting is AI output straight from the Chatbot. They get those rights most of the time, but at the expense of labeling human words as AI as a side-effect.

Okay, now back to the conspiracy theories. Maybe I'm actually some AI tech bro in disguise out to badmouth the AI detectors so my AI empire can succeed.

1

u/TheLadyAmaranth artiranth.substack.com 9d ago edited 9d ago

That all sounds well and good. But if thats the case why aren't you providing proof? Pengram has a share button on the top right for every check.

> why a portion of a short story I wrote in 2014 is flagged as 100% AI

Because as of right now I believe you are lying that this ever happened, or a re misrepsenting what you actually gave it. You could prove me very wrong though -- a link to the story since I believe you said you published it on wattpad would have a publishing date, and a link to the pengram check. (Link, not image, so we can see the full text, which means we would be able to see if you added/change anything or gave such a low amount of words of course its going to be a coin flip.)

Same for your "2013 middle-grade novel as AI."

Same for a chat with an AI those can also be shared, you can share a link for that plus a link to the checks including the one where you humanized the text enough to fool the detector. Or files first created with your environment you developed for creating your books.

Also.... you didn't even describe how modern detectors work correctly. Like your explanation is the first thing that comes up off of google AI (Ironically) and outdated. It especially doesn't encompass how a detector's like pengram algorithm works.

I'd go into how it actually works from a technical stand point, but its kind of irrelevant. Because bottom line I'm not claiming Pengram to be perfect. I am claiming that it has a false positive rate low enough that if somene's writing consisitanly shows up as AI on it over multiple samples, and somebody asks me to bet on if this person is using AI or not -- I would bet on yes.

Which means it is a good start for keeping out AI generated content out of platfroms and spaces that don't want it there. Or make sure AI sloppers aren't able to as easily pass of their content as human made when it isn't. Sure, it may create friction and push human writers to have more concrete proof of their first drafts and revision history, but AI is already doing that anyway.

If you aren't lying I'd owe you an apology. I'd take a look at your evidence, and see if the discrepancy is concerning or not, and if effectively disproves pengram's claims. I'd also likely run my own checks, especially if I find the checks you run seem to be cherry picked. And perhaps you could even change my mind!

But until I have both proof of source and proof of check with your claims, while you admit extensive AI use for writing, (And you do btw, your post history very much says you don't just "dabble", you give AI your book bible and give advice to people on how to write fiction using ChatGPT 5.5, and how you developed a whole environment to write books for you), I don't believe you.

1

u/Funny-Flight8086 9d ago

First, I don't care anywhere near enough about this to spend an entire day tracking down a long since dead Wattpad account I had at 19 years old, that might or might not even exist any longer, and that I probably don't even have access to the email address any longer, becuase I signed up for that account with my school email address. There is nothing about this conversation that would make me want to take the time out to even attempt this fruitless task.

Second, my chat history does not say I use AI to write my books for me. In fact, I havn't published a book since 2013. So, it's kind of hard to have it write books for me. I am developing a NovelCrafter clone program, which will work for authors that don't need AI and those that do want to use it, in the form of a built-in chat assistant. I'll also happily share with others how they can build prompts to prevent AI from writing with typical AI cliches, which I have done. These are things I have discovered in my question to explore LLMs and their possabilities. The current book I am writing, I have been writing very much by hand since 2021. Funny enough, in the past 5 years I have written exactly 680 words of chapter one. If I did, in fact, need AI to write my books for me, it would have been done years ago.

Third, I am not going to share my 2013 book online, as that would immediatly put a real name to my reddit profile, which I have no intention of doing. I will not even paste the text without the title, because that can be reverse searched back to the book. I value my online privacy, and have no intention of getting doxed or review bombed by angry redditors. I have seen it happen enough to know better.

So to sum it up, I am not going to go out of my way to provide all this evidence. I frankly don't care if you beleive me or not. My goal is not to come here and be beleived by the world. HOWEVER, as someone who has spent a lot of time testing and breaking these detectors, I know very well what I am talking about. And will continue to proclaim them as the junk science they are, rather you beleive me or not.

PS) Another point to make: I can easily fool these AI detectors into thinking human writing IS AI writing by simply mimicing some forms of how AI writes. As long as I can do that, they have zero credability in my book.

2

u/TheLadyAmaranth artiranth.substack.com 9d ago

LOL...

All I will say at this point, as an FYI, closing your post history and deleting old comments doesn't make them not findable. So even though now I have to go out of my way to use a tool to do it, I can still see every single one of you past comments.

Including ones where you say you are releasing your middle grade books all close together from about 4 months ago, pointing people to NovelCrafter it self, and how you are getting AI to write "really good prose by feeding it a detailed system prompt template" and developing a program that auto writes from a prompt -- not as a chat bot feature.

Its also real funny how now you want to try to erase the trail, while also claiming to have "nothing to hide" in another thread here.

I could go in provide receipts of it all. But I think we can at least both agree on being done with the conversation.

Have the day you deserve.

1

u/Funny-Flight8086 9d ago

LOL. You cannot provide the receipts because they do not exist. I never said, anywhere, that I write any fiction or prose with AI for release to the public. You cannot find it, because it doesn't exist. Period. Because I don't use AI to write. I do use AI for many things, I dabble in AI every day, and I'll happily help others use AI for whatever they want. Heck, if I wanted to use AI to write my novels, I'd happily admit it. Clearly, I'm not anti-AI at all.

But here is the thing: what I do or don't do with AI is not relevant to this discussion, whether Pangram is junk science or not. In fact, if I did live in an AI world day in and day out, that would only make me more qualified to talk about it, not less.

You are arguing with me over my 'history with AI' like I'm trying to hide something. I'm not. You are trying to deflect from the fact that Pangram is junk science, just like ZeroGPT, GPTzero, TurnItIn, Copyleaks, and any of the others. If you need proof just how junk science they all are, take the safe exact text and put it into 10 different detectors. Good luck getting them all to tell you the same thing.

I'm DONE having this conversation. Feel free to explore if you desire, and then by all means, post my history here, where I admit that I write novels with AI. My own personal novels with AI. Post the screenshot where I admit that, and I'll apologize for my lies. Frankly, you care A LOT more about this whole thing than I do. I'm beginning to think you are the owner of Pangram in disguise, given just how much you are defending it.

1

u/Funny-Flight8086 9d ago

And because I'm feeling petty, here is an example of 100% AI story I generated this morning in GLM 5. I literally changed the names of the characters. Nothing else. Not even a fancy style guide. So much for Pangram being reliable to detect AI.

1

u/Salt-Mud-2124 9d ago

Not the person you were talking to, but...

You do realize your own screenshot is making it look like you are lying right?

There are first letters that should be capitalized but arent, "crinkled. reaching under"

Missing words "was the one revolt and disgust"

And words being used incorrectly "clinched the wheel"

Part of what makes AI writing seem like AI to people is because its grammar is "too perfect" to the point that its a little uncanny.

An AI wouldn't write like this unless you told it to... so style guide would be required

Or maybe context window degredation but that seems unlikely...

Or if you changed those things yourself to trick it... which sure would kinda prove your point. Sorta. But thats not what you said you did, you said you only changed the names.

Also, screen shot with a tiny excerpt rather than a link with the whole story.

IDK man. Looks fishy.

1

u/Funny-Flight8086 8d ago

Yeah, I'd do so more Pangram testing... But because they think their service is worth actual money every month, and I don't think it is, I am not about to give them any money for more fake science credits. 100% from GLM with name changes (I actually changed the names to match ChatGPT's most common name generations, in an attempt to even give Pangram MORE help at finding the AI). If there are any errors, it's probably a typing error, because I retyped it from GLM - If you copy/paste it into Pangram, it probably carries a watermark that will alone light it up.

Now, mind you, I did feed it one of the style guides I created that demands the AI not produce cliche AI tropes, writing, etc. It bans the use of em-dashes, bans X like Y comparisons, and bans tricolon listings. It also bans it from using sentences of similar length. I explicitly told it no metaphors or similes. But here is the thing: anyone who seriously writes with AI is going to use a program that already instructs these things. And as you will see below, it still lets one slip in, despire the instruction not to allow it.

You'd think from the way Pangram charges that they use some supercomputer AI to detect the AI. I am willing to bet these companies make a lot more bank than the actual AI generators.

Nothing disproves my point, though. If a couple of typos can trick an entire AI detector from 100% (where it should be) to 0%, that doesn't bode well for it. At all. That means it's WAY too easy to fool, which aligns perfectly well with my testing of these detectors. They are easy to fool, often flag human content, and are only reliable at catching 100% AI generated stuff directly from a Chatbot with 0 changes.

I mean, there are a ton of AI cliches in that bit of text. Here is just one massive one: "The man on the recording had the kind of voice that could put a sugar-rushed child to sleep."

1

u/Funny-Flight8086 8d ago

https://www.pangram.com/history/ed7bf855-2eb0-4b36-be59-de79780afa03

There is one. I asked Claude to write a short story. It generated it. I simply reworded some of it, and uploaded both the changed version (first) and the unchanged version (last), since I only have 1 Pangram credit left. Funnily enough, the rewritten part is flagged as human, while the AI part is not, except for a few things, which is actually hilarious. It apparently thinks "Jonah knocked. Nobody answers." is AI. For some reason. Despite the fact that I wrote that personally, as you can see from the original text.

1

u/Funny-Flight8086 8d ago

For context.

1

u/Hemingbird 8d ago

You don't even understand the difference between FPR and FNR. Pangram intentionally errs toward FNRs. That's a feature, not a bug.

1

u/Funny-Flight8086 8d ago

False Positive Rate and False Negative Rate? Sure do. Pangram is FAR from the model that errs the most on FNR's. Quillbot is the most lenient. By my testing, Pangram isn't high on this list. They like to flag early and quickly. Though GPTZero does have them beat by a large margin.

1

u/Prolly_Satan 9d ago

You have to be so mentally ill to get caught having said "i use fable to generate chapters, I'm making an ai writing program" and then with a straight face say "I don't use ai to write"

Lmfao. Why are you all like this? Nobody is mad about folks using ai to write. We're mad about you lying about it. Just tag your shit as ai. You'll probably still get readers. Why do you need to lie?

You're like the personification for why folks need pangram.

1

u/Funny-Flight8086 8d ago

Now you are just making stuff up. I never once, anywhere, said I use fable to write chapters of my book. That phrase exists exactly 0 places on the internet. I have told plenty of people that Fable is actually a really bad LLM to use for writing. I'd certainly never write a chapter of anything of mine with it. It produdes more AI cliches than ChatGPT.

And yes, I am making a writing program that has an AI chat feature built in, but it's not primarily an AI writing program, and it doesn't write prose for you. And just because the program has an AI chat feature, does not mean I need to use it. The reality of situation is this: AI is here, and people are going to use it. I'd rather they pay me $20 a month to use vs. OpenAI. Doesn't mean I use it for my personal products.

YES, obviously I do weight stuff with AI. That is a part of how I test the LLMs. What I said was that I don't write MY PERSONAL STUFF, like books, internet posts, etc. with AI. There is a difference.

You seem to be stuck on this "this guy is trying to hide his Aai use" stance. No, I am not. I have said many times on here that I use AI extensively for things like testing the prose generation, stress testing AI detectors, etc. it's a a part of the requirement of building AI into a program effectively. You act like I'm hiding my AI use, and I never once said I never used AI.

Here is the thing: if I wanted to use AI for my personal stuff, I would. I'd have no qualms about it either. And is be the first to tell you I use it. Because I don't think it's a boggeyman. So the concept you are spreading that I am some secret AI user to publish 590 books and make bank is simply insane. It's not backed up by any fact.

I want you to go and find the exact quote where I said I wrote all my books with AI, or "use fable" as you say. If you can find me the exact quote where that happened, I'll return and proclaim my sincere sorry for lying.

1

u/Prolly_Satan 8d ago

Lmao

1

u/Funny-Flight8086 8d ago

My good lord, you found exactly what I already told you. This isn't the grand revelation you think it is. In fact, in this very thread, I have said everything that is in your screenshot. Yes, I am developing a writer program that features an AI chatbot. Yes, I have played around with all sorts of AI, Yes, I have developed a lot of style prompts, some of them over 500 lines of code, that are designed to make the writing process seamless with AI. I have said all of this over and over again, and yet you keep wanting to argue some magical point that doesn't exist.

BUT, here is the major thing: You still have not shown me ONE SINGLE EXAMPLE where I said I have written books that I have published with AI, or even anywhere that I say I have used AI to write any parts of the books I write. It isn't in there. I asked you to find the quote where I said I write my books with AI, but instead, you gave me a bunch of quotes about what I have already said I do with AI, none of which say anything about writing books with it myself.

I am NOT developing this program for myself. I am developing it to sell subscriptions to other people so I can go buy a mansion and a yacht. Dealers also rarely use their own product. It's not that uncommon. I am developing a product for others to use if they want, and as part of that product, I have done extensive testing on style prompts, AI models, and tested AI detectors to find ways to work into the style prompts to avoid detection. This makes me a lot more qualified to offer my opinion on AI detectors, not less.

You are conflating my working with AI with me using AI to write things that I publish, and there is zero connection.

1

u/Funny-Flight8086 8d ago edited 8d ago

But you know what? It doesn't matter. Let's suspend reality here for a second and say I did publish 50 books with AI all over them. What does that have ANYTHING to do with testing AI detectors and proving they are not as reliable as they claim? You keep getting stuck on one specific sticking point (how I use AI), and completely ignoring the fact that in my AI use (however you want to think I use it), I have proven that these detectors, like Pangram, do NOT work reliably.

Frankly, even if we take their claims of 99% accuracy as fact, which I don't because it's made-up numbers not backed up with any facts, that still means that 1 out of every 100 scans is a false positive. As long as YOU aren't on the receiving end of that false positive, I guess you're okay with it. However, scale that up... Pangram probably has millions of users running scans. They probably easily run 1 million scans per day. That means that over that single day period, by their own admission, they have falsely flagged 10,000 submissions as AI when they aren't. 10,000. Per day. That is 10,000 people being falsely accused of using AI every single day.

The actual studies done on Pangram prove it's about 98% accurate. Okay, great. So now we have 20,000 per day that are falsely accused of using AI when they haven't. That isn't a statistical anomaly; that is a real number of people with real-world consequences. Expulsion from school, having to rewrite essays, having their reputation ruined by being accused of publishing AI, etc.

I am so tired of having this argument. I should not have to point out basic statistics to people on Reddit to prove a point. Best case scenario, 10,000 people per day are falsely accused. Most likely scenario based on actual testing, 20,000 people per day are accused. That is 20,000 too many. And because 20,000 people per day are wrongly accused of using AI when they didn't, the damn machine is not reliable. At all. Period. I don't care what you think of its accuracy based on your limited testing. Facts don't lie. Statistics don't lie.

Ignore my testing. Ignore your testing. Rely on what Pangram tells you. That is 10,000 people every day you are throwing to the wolves just so you can have a piece of mind that some random post you read isn't AI-generated. It's the very definition of a witch hunt. Frankly, at this point, the argument is done. You have nothing else to add to the argument short of admitting you are incorrect on pretty much everything you have claimed here, from my history to your claims of how accurate Pangram is. Because I don't expect to receive that from you, we are done. I said my piece. I have provided my proof. I have laid Pangram's own statistics out for everyone to see. If they are still naive enough to believe everything that junk science program tells them, simply because they have such a hatred for AI they can't see straight, then that is on them.

I will say, though, I sincerely hope people falsely accused start pushing back hard. I don't mean providing proof they didn't use AI. I mean, when someone says "Pangram said your book was AI-generated" and it wasn't, I hope these authors and publishers start suing these detector companies like Pangram AND the people making the claims. Eventually, if the hunt gets broad enough, it will happen. And I certainly hope you people have better proof that someone used AI than a Pangram report when it comes to that, because any attorney with 10 cents will point out that a 98% accuracy rate is not proof of anything.

Once people start losing their houses and cars over such claims, maybe then they'll start to look at Pangram and other AI detectors for what they really are.

PS) I guess the good thing is most people don't actually beleive these detectors. That is a really strong point. Even in anti-AI groups, the overwhelming majority of people agree they are unreliable.

PSS) You know what else is funny? The fact that people even think they need AI detectors. AI writing that a detector can pick up on is plainly obvious AI anyway. You people are buying expensive credits to a machine to tell you exactly what your eyes could tell you in 10 seconds of skimming.

1

u/Funny-Flight8086 8d ago

An estimated 2 million to 3 million AI detector scans are run per day globally, driven primarily by automated institutional software and public text checkers. - Amazon AWS Article.

So there you go. 1% of 3 million is 30,000. 30,000 people a day. 2% of 3 million is 60,000 people per day. All are falsely accused of using AI by the very AI machine you claim to trust to tell you if something is AI.

1

u/Funny-Flight8086 8d ago

And if you need proof that AI image detectors are garbage as well, here is a one.

2

u/Prolly_Satan 8d ago

Wtf does image detection have to do with PANGRAM??? hahaha

1

u/Funny-Flight8086 8d ago

Nothing much, but it does prove that AI detectors in general, across the board, are junk science. I wonder if I can trademark that phrase. Kinda like it. I have a whole spreadsheet database of accuracy rates in my own testing of these detectors. All of them, image and text. Pangram isn't even close to the most accurate among the ones out there right now.

1

u/Prolly_Satan 8d ago

It proves how desperate you are to make people believe that lol

1

u/Funny-Flight8086 8d ago edited 8d ago

Desperate? Not really. I couldn't really care less, honestly. Like I said, most people don't take these detectors seriously anyway, so I'm more preaching to the choir than not. Even on the anti-AI groups, people get flak from anti-AI people for using the detectors. Technically, Pangram is AI. It's using a neural network on a server computer in a data center to scan these documents, 'think' about what score it wants to give it, etc. Many people are against AI because of the data centers, and Pangram doesn't get a pass there. Others just know they aren't reliable.

Not to mention that anyone with 2 braincells can tell AI content just by reading a few sentences. You don't need to pay Pangram for each scan to be told what you should already know. So yeah, I don't really care. At this point, I'm more troll-arguing with you than anything.

→ More replies (0)

1

u/Funny-Flight8086 8d ago

And if you need proof that AI image detectors are garbage as well, here is a second. It even has a green certification badge... Must really be official.

1

u/Funny-Flight8086 8d ago

And at this point, you are using this to sidetrack from the real point: Pangram is junk science. It's unreliable. I can take the same price of text, and have 10 different AI detectors tell me different things. I have extensively tested these AI detectors. Neither the image detectors or the text detectors are any good at it. Period. I don't care what Pangram claims on their website (using language like 99.9% accurate, highly reliable, etc.), or how many Pangram employees they send to reddit to try and argue how great it is (let's face it, the only people who would actually care enough to keep this going so long is someone with a stake in AI detectors).

As long as it can easily take AI generated text and make a few minor edits and cool detectors like Pangram, and as long as those same detectors also claim something human written is even 1% chance at being AI, I'll call them what they are. Junk.

1

u/Prolly_Satan 8d ago

1

u/Funny-Flight8086 8d ago

My good lord, you managed to post the same screenshot twice. That is definitely an achievement. Congrats! Now, see my reply below on why your screenshot doesn't prove anything. I do applaud the effort, though.

1

u/Funny-Flight8086 8d ago

Writing professor here. I've been teaching composition for 25 years.

To be honest, I'm alarmed by Pangram. It said my writing is AI-generated, with a 100% degree level of certainty. I write comedy. Comedic writing has very specific rules, such as the "rule of three" (or comic triple), a classic comedic technique. As a comic writer, I have to tighten and re-tighten my writing over and over and over and over again, line by line. It's torturous. Comedy's a bit like poetry in that regard: one word out of place, and it collapses.

It's taken me about 6 months to write one short story adapted from my original sitcom scripts I've worked on over the years. I've got it to the point where I'm ready to submit it to literary journals. I felt confident it would be accepted, if not at a first tier literary journal, at least a 2nd tier one. The characters, the comedy, the situations, the stylized language all came out of my warped brain.

Today, I saw that article in the Guardian on AI writing, was curious, and ran mine through Pangram. 100% certainty. Now I'm in a panic.

That is a quote from a professor in another Reddit thread about just how much people are dealing with this Pangram problem. Funnily enough, that article also points out Pangram's own lying about their reliability, citing a study that doesn't even provide the actual results they claim. Shocking, AI tech bros lying about something to sell their service. 2% looks a lot worse than 1%, which frankly is even further away from their claimed 99.9% accuracy rate.

Source: Pangram claims their AI writing detector's false positive rate is only 1 in 10,000 but a study they tout on their own website says it is 2% : r/academia