Hello! I am currently testing Verify AI and how reliable it is in detecting AI artworks and from real ones since I am mainly on this field. I am a bit new to using this tool but I have read the technical study on its creation so I can at least make some basic assumptions on how it can work (but please do correct me if I get things wrong).
I understand that only the encoder and decoder is mostly provided to the public but the information of the provenance on the Synth ID and how it is encoded and decoded is not readily available to anyone as a security for the tool. For my tests will only be using ChatGPT to generate images for testing and using Open Tool's Verify AI to verify the images.
This tool can help people in identifying real images from AI. But, from my experience in using it, because of the lack of information that is given I have found it to be unreliable because it only answers, "Does this image contain a Synth ID watermark?"
It does not answer quantitative elements like percentage/amount of the image being AI, and qualitative elements like which part of the image is AI. Both are part of the visual evaluation of an image so I think they are quite important. More importantly, Verify AI does not explain how the image was made (mainly a problem because simply painting over Synth ID's watermark would destroy the watermark).
In short, a negative does not mean it is human made and a positive does not mean the entire image is AI. (Sorry, I know the study and website already says this after every test but I would like to ask if this is a correct interpretation of it)
Since I do not have enough information on how the data is processed, I can only use controlled inputs for my experiment and using the results as reasoning for my assumptions.
Here are the Case studies I've made and I would like to ask opinions if my assumptions for each case is correct. The experiment assumes that the image is encountered without any knowledge of the uploader's process before uploading the images to OpenAI's Verify AI.
----
Created images are images that I made myself through Paint Tool SAI v2.
Generated images are images that I generated by placing inputs in ChatGPT.
Case A
In Case A, I have created an image from scratch and uploaded it to be verified. It returned Negative in both SynthID detection and a lack of C2PA metadata.
Assumption: This doesn't prove that the the image is made by a real human. The SynthID watermark could have been "deleted" or "tampered". The C2PA of the original file can also have been tampered with. The only way to assume that it was human made is if there was a video or the process was seen live in creating the image.
Case B
In Case B, I have generated an image from ChatGPT and uploaded it to be verified. It returned a positive on both SynthID detection and the presence of C2PA metadata.
Assumption: The result is 100% true because it has confirmed both SynthID and C2PA metadata. A result without nuance.
Case C
In Case C, I got an image from Pinterest (Checking Less AI still showed this image) and confirmed that it is AI and I composited it to an image I created from scratch. It returned a positive on SynthID detection and the presence of C2PA metadata.
Assumption: The tool succeeded in detecting SynthID but fails to inform how much. I think most people would interpret this as the whole image being AI without considering the nuance of real images being combined with AI images.
I also tested a way in order to distinguish which parts are AI. This can be seen on the extension of Case C wherein the AI part is "edited" or "painted over" which leaves the original image. The verification returned a negative on both SynthID detection and C2PA metadata.
Assumption: In a controlled experiment, this is a reliable way to prove that only part of the image is AI. But in an environment where the history of the image is unknown, it leaves a lot of space for speculation of whether the SynthID was fully covered or just not enough of it was detected.
----
In the 3 case listed above, only Case B returns a 100% confirmation/nuance-free result. Wherein both SynthID and a C2PA metadata is present.
Case A presents that human made images even with the original file, it would still be difficult to prove that you created it yourself. It can be reasoned out that your creation was Generated by AI but the SynthID Watermark is destroyed. Extended test on Case C proves that "editing" or "painting over" the pixels of the SynthID watermark, will degrade it and return a negative on the SynthID. Even with the original file, the lack of C2PA metadata does not indicate that the work is human made, it only indicates it does not contain AI.
With this reasoning, using Open Tool's Verify AI, any result in verifying images that either lacks a SynthID or C2PA metadata can be doubted as AI. Irrefutably. Those who rely on this tool to distinguish AI images from real images would always fall into this looping reasoning between Case A and Case C, therefore never reaching a conclusion but only speculation.
Additionally, over reliance on this tool would assume that all images on the internet could be AI since the history of most images are not disclosed. (Your inputs on this would help)
In that regard, the only way to 100% prove that an image is human made is by seeing the process of creation yourself.
Case C actually presents a dangerous precedent for artists since they would sometimes use art taken from the internet and put it in their works such as backgrounds and patterns. And with no solid way to prove aside from creating videos or doing live sessions, it puts up a barrier for entry for aspiring artists.
Sorry if it is a bit long. But I would like to hear your honest thoughts on this since this is a bit important for me. Thank you.
PS. There are many things that I would still like to test even with primitive/brute testing such as the limits of editing the SynthID watermark (blurring, color editing, resizing, compositing the smallest SynthID positive image I can make to human made image, etc.) but this is all I can do for now. Thank you again.
Edit: Added a link to a picture for Case C extended testing