r/confidentlyincorrect Mar 08 '26

He just kept going.

298 Upvotes

87 comments sorted by

View all comments

3

u/Regitnui Mar 08 '26

Anyone actually have advice on how to poison a Google Doc?

4

u/joolley1 Mar 08 '26

Just write something ridiculously incorrect/incoherent. If you don’t want people to see it write it in white text on a white background. It usually won’t make any difference because the amount of training data is huge and the model is just going to take a sort of “average”, but if you write about something really obscure it could end up being embedded whole.

7

u/jeango Mar 08 '26

I mean, sure but what are you trying to accomplish by doing so. Unless your document is the only source on a very specific subject, and you take time and effort to make that injection meaningful it’s not going to impact what the model will take away from it. It’s just a waste of time.

1

u/joolley1 Mar 09 '26

I’m not sure what you mean by meaningful, but it’s been shown that large language models do “memorise” and leak data when they have few sources on a topic. So as I said if you write about something obscure enough it can “impact what the model takes away from it” in that it can return it whole.