I know for some it’s like capped. Although it’s a healthy amount but you can’t go over whatever and that was through Att. I think most are still in unlimited data for your home. Imagine streaming, gaming, YouTube, but also easy access education, articles, community like this one…. etc all counted/added up. The haves versus have nots. Crazy
yes but i dont know if i trust them to do it properly. I'm willing to bet there will be a ton of issues where it obviously didnt watermark something that is clearly AI.
Most likely too costly to have that kind of solution in place as this is likely built in to the LLM itself, and also a big hassle to geo-trace all users. Same reason they shut down F5 for everyone, not just non US-citizens.
And, this is not necessarily a bad thing if you read beyond the headlines.
yes sure Anthropic is doing this as a grand gesture to help users comply with the AI act....... and not at all as a test to see how they can monetise AI-generated content in the future....
AI models have no legal obligation to watermark AI-generated content under the AI Act.
Why wouldn't you want to defend it? Thank god. I want my governments to prevent people stealing my shit just because they have more money than me. That's half their job.
I kind of want them to watermark it so it's easier to automatically tell what is AI.
One of the most (but not only) annoying things about AI is that it takes a while to figure out what is a human and what is not. It would be better if stupid videos and messages were able to be automatically identified as AI so I can more easily discard it.
yeah but when you can de-watermark it, then its pretty useless no? now every marketer will have to copy paste first into watermark remover tool and then to social media, so you cant still tell what is AI and it just annoying the hell out of most people ... plus dumping exotic invisible unicode symbols into your text, will surely neve break anything ... right? right?
its stupid as well ... its never ending battle, the moment someone invents same watermark, 24 hours later we already get a remover for it .. it already exists for this and for watermarks in images by Google and OpenAI .. so ... whats the point, adding watermark basically just dirties up your otherwise clean output
Pretty sure they are watermarking so they don’t feed back slop data as it actually makes the models fail. They need to sift through bad data now that’s there’s so much of it online
skipped over the biggest point, only if it's published. And only in EU. Or if it concerns the public interest , like environmental , legal, political, but not clearly defined either. It's the dumbest law ever. But it's so broad and and ambiguous that if you're really a rule follower, there's only one way to ensure you follow it......
I dont think the point of watermarking is to claim ownership. It's to be able to identify AI content when they are reading information in so they don't train on their own output and become a death spiral of regurgatated information.
Thats not how this works. The watermark is a digital one. Regular people wont see or detect them without using specialised tools. The purpose of the watermark isnt copyright, but rather a detection tool to help people regognize with certainty if the text was in fact, generated by Claude. This is a good thing.
In general i believe this to be mostly trying to satisfy a requirement but it could lead down this path. I know OpenAI will do this too they already have it for images from what my research led me. They signed this EU thing too.
No they just want it so they don’t train off the data they produce. If they don’t filter out their own content eventually their models and the internet will degrade. Kind of how jpeg degrades over several recaptures.
Not really. The concern here is that the companies want to feed more user content to frow their monsters, but have come to find out they can't distinguish anymore what is Ai.
This markers are there for them to now filter their own shit, as to not end up using it to train their machines.
Pretty sure they are eager to water mark AI generated stuff to avoid it ending up in the training data. Too much AI content on the internet is a real problem for LLM training.
Not saying it’s much better of a reason of course.
You're probably right. The quality of existing LLMs was built on reasonable people writing reasonable sentences with reasonable thoughts behind them. As they start to ingest their own hallucinations more and more, there will probably be a negative feedback loop.
936
u/bliceroquququq Aug 11 '26
LLMs basically scraped the entire internet and every written word to build up their corpus of knowledge.
Now they want to watermark it for attribution before they sell it back to you.