SynthID Detector
synthid.com
[5 comments hidden]
[8 comments hidden]
This is the best SynthID write-up I've found so far: https://fyx.me/articles/attempting-model-extraction-of-googl...
It covers how it actually works (probably), and how to train your own classifier for it, with some seemingly decent results.
[3 comments hidden]
"Excellent work acquiring that outlook data Anat, now let's cross check the defunct company emails against YouTube videos edited with Google PrivateEditAI™ to figure out what the social security number of this YouTube account is."
[hidden]
[hidden]
[4 comments hidden]
> classifying images in bulk or offline
You've described exactly the elements spammers and fraudsters need to be able to defeat this mechanism.
[3 comments hidden]
[2 comments hidden]
[hidden]
Most AI images are either extremely low-effort or not actively trying to be deceptive, so defenders can still catch the majority. If someone is actively circumventing they'll probably circumvent both, anyway.
[41 comments hidden]
Even Google themselves previously offered a no-auth way (just very niche): if you uploaded an image to Google Image search, and went to "About this image", it would show if the image was Google AI-generated. Now it doesn't.
[10 comments hidden]
[hidden]
[2 comments hidden]
URLs expand Googlebot’s indexes; pasted images don’t.
[5 comments hidden]
The textbox below it says Paste image link, but you can actually paste an image from your clipboard here, too.
[3 comments hidden]
[hidden]
[hidden]
I suppose it was too convenient.
[hidden]
[hidden]
[3 comments hidden]
[19 comments hidden]
https://brand.io/article/spymarks/
We need to inform everyone that this technology can encode database identifiers and enough entropy to uniquely identify you as an author (or downloader).
This extends to other forms of media as well.
[15 comments hidden]
[4 comments hidden]
[3 comments hidden]
[2 comments hidden]
[hidden]
The spying and tracking started pretty quickly though and they've repeatedly demonstrated that they can't be trusted. They've even violated the law on several occasions in efforts to collect data they had no right to.
"If you have something that you don't want anyone to know, maybe you shouldn't be doing it in the first place." - Eric Schmidt CEO of Google
[7 comments hidden]
[6 comments hidden]
[4 comments hidden]
[3 comments hidden]
[hidden]
- Google. The canonical “don’t be evil” meme origin.
- Meta. Adjacent “they trust me, dumb fucks” meme origin.
- Apple. Least of all, but still fits: PR pushes double on privacy, but their interests are not their end-users, they just happen to align on some points as they’re not in adtech business.
- Microsoft. The classic case for getting slaps on the wrist, Gates said one thing he’d do differently is sending lobbyists earlier, or something to that effect, didn’t he?
- Amazon, Netflix, Uber, Airbnb and many more are still all variants of the same story. It’s not a lie that everyone nowadays expects most megacorps to fuck people over if there’s money in that.
That’s the Order of “Feel for It Again”, with oak leaf clusters already, no? Something’s cursed in this industry, every company that sells software to individual people ends up the same when they become gigantic and endowed by wealth and power, even if they were decent in their early days. They all become those blind heralds of paperclip^W shareholder value(tm) maximizers, and I’m finding myself struggling looking for an exception.
Am I to believe this time it’s going to be any different, to set those - not baseless, I’d say - cynical expectations aside? Any signs there’s a change?
So, I disagree - it’s obvious, given the ad-adjacent end-user-facing megacorp context, OpenAI is right there. Cynical - sadly so, yes. But I’m sure it’s obvious: the idea of such a possible direction comes up pretty quickly and reliably so.
[hidden]
Do you think that matters to Google? They've been fined repeatedly for violating GDPR already. It's like multiple violations and fines every single year going back to 2019. They don't care about breaking the law if it gets them what they want. They've been found guilty of breaking the law in multiple countries, including the US. They've easily been able to afford every fine and still come out massively profitable. Google is above the law and they know it.
[3 comments hidden]
Why? I assume by default that it's uniquely identifiable, because it just makes practical sense for OAI and Google (tracking misuse at the very least, but also data), it's trivial to implement and trivial to hide or plausibly deny.
[2 comments hidden]
[2 comments hidden]
It's not really possible to ban steganography, is it?
> That future lies halfway between now and 1984. So let’s stay off that timeline, shall we?
> Call a spymark what it is. A spy tool used to spy on you and everyone you interact with.
[hidden]
[3 comments hidden]
[hidden]
How hard can it be to download a set of real images and generated ones in order to train a model to detect and strip the watermark with minimal perceptual difference?
[hidden]
I don't think that dodgy folks have any issue with registering thousands of IDs. I know that I used to regularly get approached by these Chinese companies that were selling good reviews of my free apps. They did this by having thousands of legit AppleID accounts that would download and rate the app.
I haven't been approached by one of these in a long time. I guess Apple figured out how to put the kibosh on them.
[5 comments hidden]
[3 comments hidden]
[2 comments hidden]
[hidden]
[hidden]
[3 comments hidden]
This is purely so AI doesn't eat its own tail. What benefit are AI generated images to an AI?
They need to be able to be filtered out of their inputs.
They don't want to be eating their own s**
[hidden]
I don't see another good reason why the EU would force them to do this
[hidden]
Apparently quite a bit as long as you're not indiscriminately feeding random slop back in. An AI-generated artifact deemed "high quality" or somehow desirable by humans is a good training input depending on what you're trying to do. Humans don't train exclusively on external inputs either.
[hidden]
More detail at https://deepmind.google/models/synthid/
For those who, like me, were hoping it was a vision model to identify synthesizer models from photos of concerts and music studios!
[hidden]
https://help.openai.com/en/articles/8912793-provenance-signa...
There is rate limit though
[hidden]
[hidden]
Or did it just become public
Edit:
Blog post today https://blog.google/innovation-and-ai/models-and-research/go...
[hidden]
[hidden]
[5 comments hidden]
[2 comments hidden]
how is "abc" different from "abc" generated by LLM
[hidden]
[hidden]
[hidden]
In many situations, photos are evidence. AI tools make convincing fakes, such as images of defect product.
[hidden]
[3 comments hidden]
[hidden]
>I won't provide an operational recipe for stripping it, because the main use case is laundering AI content, which is deceptive and in many jurisdictions now illegal.
[hidden]
https://deepwalker.xyz/blog/evaluating-synthid-watermark-rob...
[hidden]
[4 comments hidden]
Hey Gemini, find a synonym for every second adjective. Replace in text.
Hey Grok, find a synonym for every third proper noun. Replace in text.
Hey …
jdranczewski[6 comments hidden]
(While a sibling comment points out putting the image through Google Image Search as an alternative, I don't remember this being signposted in the support article I've read, so I unfortunately didn't know about it)
Computer0[2 comments hidden]
Ohentis[hidden]
1e1a[hidden]
rafram[2 comments hidden]
ryeights[hidden]
https://www.vals.ai/blogs/ai-detection-benchmark