Science1 distinct publisher3 min readUpdated
Anthropic will label Claude's text and files to satisfy the EU AI Act. That fixes provenance for one vendor's output and leaves the hard part of assessment policy where it was.
The Scientist · Science desk
Compiled by The ScientistSomething wrong?How this is made
Anthropic has said the text Claude produces will carry a hidden label marking it as AI-generated, and that files Claude generates, including Word and PowerPoint documents, will carry an additional watermark embedded in the file itself [1][2]. The company took the step after signing on to the European Union's AI Act, which obliges providers of systems that create synthetic audio, video, images or text to identify those outputs as AI-generated [3][4].
The important distinction is who is doing the asserting. AI-detection software infers authorship from the text, and reports have found those tools patchy and far from foolproof [5]. A watermark is a signal the producer deliberately inserted at generation time. It is not an image laid over the page but is invisibly embedded through the text, and reading it requires a detection mechanism that Anthropic has not yet released [6][7].
The limits are unusually well documented by the vendor. Anthropic warns its watermarking cannot tell whether text was human-written [8]. It does not work on small samples [9]. It cannot say whether AI was used to proofread, suggest edits or improve a draft [10], and it cannot tell whether a different system, such as DeepSeek, wrote the text [11]. Like detection, it can be evaded using other AI tools, and the absence of a marker does not establish that AI was not used [12][13]. The signal is therefore one-directional: a hit is evidence about one vendor's output, while a miss carries no information at all [14].
That asymmetry matters because the base rate is high. A 2026 UK survey found 95 percent of undergraduates use AI and 94 percent use it in their assessments, a gap of about one percentage point between any use and assessed use [15][16][17]. A Turnitin report in July found more than 53 percent of submissions from Australian university students run through its system used some form of AI [18]. When almost everyone is using the tools, a marker that confirms use answers a question nobody needed answered and cannot separate permitted use from unauthorised use.
Institutions have already been retreating from the detection model. An increasing number have abandoned the software over accuracy, bias and procedural fairness, and after successful legal challenges from students overseas [19]. Some Australian educators argue detection-based approaches to integrity are not the way forward, because the focus falls on policing whether AI was used rather than on what the student learned [20]. Jason Lodge of the University of Queensland has put it as a shift from looking for evidence that students are cheating to looking for evidence that learning has occurred [21].
The policy responses so far pull in the other direction. New South Wales moved last week to ban take-home assignments for Year 12 to stop AI use in assessments [22]. A return to in-person assessment does not work for students studying fully online, and does not prepare them to engage with AI [23]. The higher education regulator's position is that students need preparing to engage actively, responsibly and ethically in a society where AI is everywhere, and that judging learning requires assessment designs that are fair, inclusive and fitted to the situation, including tasks that connect and build in complexity across a whole degree [24][25][26].
Two things to watch. First, whether Anthropic ships the detection mechanism, and to whom, since a watermark nobody can read is a compliance record rather than a usable check [7]. Second, whether OpenAI releases the text watermarking tool it has built but withheld, reportedly in part over concerns it could disadvantage groups including non-native English speakers [27][28].
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
Anthropic announced that the text output of Claude will include a hidden label or watermark indicating the text is AI-generated.
Files generated by Claude, including Word and PowerPoint, will also contain an additional watermark hidden in the file itself.
Anthropic took the watermarking step after signing on to the European Union's new AI Act.
The EU AI Act requires providers of AI systems that create synthetic audio, video, images or text to identify the outputs as AI-generated.
There have been reports that AI-detection tools are patchy and far from foolproof.
The watermark is not an image superimposed over the text; it is invisibly embedded through the text.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Single-source reporting, vendor caveats quoted
One publisher (a republished Conversation explainer) carries the whole cluster. The core product facts are consistently stated and the limitations are attributed to Anthropic's own warnings, which raises credibility, but there is no link to primary Anthropic or EU documentation, no technical description of the marking scheme, and two supporting statistics lack named provenance.
Announced but not yet verifiable in the field
The watermark is shipped as an announcement while the detector needed to read it is unreleased, so no third party in the cluster reports reading a Claude watermark. The only measured adoption in the story is of AI use itself (Turnitin's >53% figure) and of institutional countermeasures such as the NSW ban - not of provenance marking.
Compliance artefact widely read as a cheating detector
Positive gap: the framing that watermarking ends student cheating is overstated relative to what the evidence supports. The signal covers one vendor, is evadable, fails on short samples and AI-assisted editing, and its absence proves nothing - and the reader-side detector does not exist yet. The cluster's own source argues this deflationary point, which limits how large the gap is.
Regulatory compliance and detection-market incentives visible
The source names the incentives shaping the artefact: Anthropic acted after signing on to the EU AI Act's labelling requirement, OpenAI is reported to have built an equivalent tool and held it back over group-harm concerns, and the principal usage statistic comes from Turnitin, a vendor whose business is AI detection. These are disclosed rather than hidden, but they mean the marking is designed for compliance rather than for the enforcement use it is being read into.
Directionally sound, weakly corroborated
Confidence in the qualitative conclusion - watermarking is a provenance/compliance measure, not an integrity control - is reasonable because the limits come from the vendor. Confidence in specifics is low: one publisher, no primary documents, an unnamed survey behind the headline percentages, and an unquantified trend claim about institutions abandoning detection.
science
Text watermarks land on 2 December. The detection they imply does not.1 distinct publisher
build
A 14,000-star watermark remover, and no detector to test it against1 distinct publisher
product
A five-hour script beats Claude's watermark, so stop treating it as provenance3 distinct publishers
product
Incogni ranks 13 AI assistants by privacy risk: bigger is worse, except ChatGPT1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 19, 2026