I thought declaude would be like degoogle.
Nope, turns out what they do is precisely the opposite. They try to remove those described markers. Real scummy.
This is a most excellent place for technology news and articles.
I thought declaude would be like degoogle.
Nope, turns out what they do is precisely the opposite. They try to remove those described markers. Real scummy.
Ah, that explains why the article smells like an LLM wrote it.
I think the thing that frustrates me most about AI is that so many people seem to have forgotten that writers exist and wrote huge amounts of text - books, news articles, magazines - for centuries. But now, anytime someone writes a few cogent paragraphs, you get people screaming "AI!!! IT'S AI!!!!!!!".
I know, because it has happened to me.
The article, to me, sounds like a competent author wrote it. Which is representative of a lot of the text LLMs were trained on.
I can't tell you about the text, but the whole site is vibe coded. The CSS, the js, everything. Which doesn't really inspire much confidence on the text, especially as it's explaining how to bypass the claude detector.
I'm against trying to "censor" LLMs, but yeah. What's even the ostensible benevolent scenario for stripping invisible watermarking from Claude?
I can't think of one, playing devil's advocate.
It's pretty scummy.
I am against AI and LLMs but the reason is quite simple, a lot of modern writing is offical bullshit to get stuff approved. Research grants, medical therapy approval, offical work mail that has to sound professional, job applications (insofar that noone cares what you write, they want a standard template and pick a candidate along qualifications anyways.). All these texts cost a lot of time and it makes no difference if a human or some copy machine writes it. And tbh. were it not for all the disadvantages of modern LLMs i would use it for the exact same reason, because sitting at a grant application for two hours while others do it in 5min is fucking useless and exhausting.
If it's stuff that no one cares what you write, then they also won't care if it's proveably written by AI or not. So removing the watermark still does nothing beneficial.
Assuming a checker tool is public it's enough for a teacher to verify classwork, but AI providers having the sole ability to identify generated content with no way to independently verify is not a real solution.
Plus this site advertises a tool to remove this watermarking, so it can't be that hard to scrub out if you're aware of it.
The goal is to ensure that they don’t inbreed their models, not fix the problems they’ve caused.
This won't fix the inbreeding issue anyway. The bias is extremely slight, but random, and orthogonal to Claude's own "slop patterns" and tendencies.
To put it another way, its way below the noise floor of output at ~1 temperature, anyway.
I wonder what would happen if some start to watermark document they doesn't want in Claude training
Very interesting read.
Well shit, that was fuckin’ interesting.
Yes, AI is evil. So is facebook. But the engineering is still interesting.
That's one of the things that frustrates me most about this whole AI thing. I fucking hate it and I want it to die, I wish it were never created in the first place. But from a tech enthusiast and a maths nerd point of view, it is super interesting.
Like the performance of these models is shit compared to a real person doing actual work. But if we think about what we are doing on a basic level, the performance is way beyond what I would expect it to be. I wouldn't expect it to be able to form a coherent sentence or scale as well as it does (even though the resources required to run these is still very high).
It could have been really cool shit people did studies on and played around with to explore the math. Cool little play models we could let go on a bunch of data and see what it did and how. Something for a small group of nerds and experts who are into that kind of thing, for the sake of learning and nothing else.
But no, somehow it got transmorphed into "AI". And marketed like this actual learning almost sentient computer system that can replace all workers. You can ask it anything and it will give PhD level expert answers. Oh and it's run by a handful of the most vile men imaginable who pour all of the world's money and resources into it, all so they get to be god emperor of the world. Fucking terrible.
The emergent behavior in huge models where it can “reason” instead of simply predicting the next token is fascinating. Artificial “neurons” built on statistics and linear algebra emerge to create something that legitimately has artificial intelligence. ANNs were conceived in the 1940s, building on centuries of development in statistical modeling and only now do we have the compute power to make this vision a reality.
Yes, the “intelligence” has significant limitations and won’t be replacing human intelligence anytime soon, but it can actually be a useful tool if its limitations are kept in mind.
The problem is of course the tech bros turning centuries of innovation they had no part in developing into a massive Ponzi scheme for their own profit.
AI isn't evil. Generative AI isn't evil. AI has existed for 20+ years now, I studied it back in my uni days.
Corporations, how they trained it, how they use it now, how they are willing to pave the planet to force it down our throats is evil.
This is one of those things as tech people we have to come to terms with and understand. No technology is inherently good or evil, it's what people do with it.
I wonder how many people will be identified as AI because they used AI so much they started constructing sentences like AI.
This particular watermarking would be effectively impossible for a person to end up replicating.
As a neurodivergent (possibly AuDHD), I wonder and worry if this wouldn't end up increasing false positives when it comes to neurodivergent way of speaking.
It won't, unless you normally talk almost exactly like the model in question and then alter your word choice distribution according to the specific secret key entropy.
I'm not convinced that they even know 100% how Anthropic is doing it. I can think of an easier way that doesn't corrupt the text: just find a bunch of tokens where there is a good spread of token possibilities, and the more often the most likely one is chosen, the more likely it's AI.
That being said, it doesn't seem much different from what any of us do to identify AI text — it has lots of tells anyway.
That's how AI testers work and its why they don't. Most forms of formal writing are predictable by design. If the AI can predict predictable formulaic writing, it doesn't mean its AI, its probably just any form of professional writing other than fiction.
Famous public domain works will always be considered AI by those tests, because of course your LLM knows the american national constitution. It was in the training data, so it can predict it with 100% accuracy, therefore your test wrongly calls it AI.
Testing for AI writing that way does not work.
Interesting stuff. My own far less scientific reading of the article itself seems to fittingly suggest it too is largely if not entirely AI generated, which I guess would make sense.
95% AI text agree. It reeks.
Although this looks like a clever approach, a kind of stochastic key, I do not see how this guarantees to distinguish text written by big babble machines versus humans. Humans also have a certain pattern of writing, a given distribution of how some words are more likely to appear than others. How can one tell them really apart?
As an indicator, yeah, might be usable. But I wouldn't read too much into it before seeing results of a study that runs actual tests.
I thought the article explained that pretty reasonably on a scale of probability and weight. The longer the text, the more reliable the scoring.
It's not about the variation of the words, it's about the variation of the words from the model baseline.
Like if your word choice was almost the exact same as Claude's normally, maybe you just talked to them a lot and picked up their phrases like it's not nothing.
But if you managed to be almost exactly like Claude and yet varied the possible words exactly according to a hidden entropy key, they'd know it was actually Claude with the SymthID-Text watermarking applied, as no human would end up falling into that statistical bucket.
- Only the key-holder can check. Your teacher, editor, or favourite "AI detector" website cannot run this test; a genuine check needs the provider's secret key, or a checking service the provider runs. Google runs an early-access detector portal for SynthID; Anthropic says detection tooling is forthcoming.
I am not so sure about that. The amounts of words is finite and with enough text, you will see that certain words are used more often, especially in certain combinations. I believe people will brute force this and then create a way to destroy the watermark again.
I wonder if they did this to appease the EU or just to have a way to prove in court that a specific competitor distilled their model using claude
Sell access to Turnitin and the likes for a small fortune. They are all but required to pay whatever the price is.
They're invisible, they survive copying, and they work because they don't live in the characters at all. They live in the choices between words.
So it is yet another way for AI to falsely hurt neurodivergent people over their writing styles and word choices.
Fuck AI.
I was not sure how any of this worked, and those interactive demos along with the explanation are quite helpful.
Also a very important point made about this not being a generic AI detector at all, and only being available to the model creator.