Invisible Watermark Added to Claude’s Text Output by Anthropic

Invisible Watermark Added to Claude's Text Output by Anthropic

Anthropic Introduces Watermarking for AI Text Generation

On Monday, Anthropic revealed that its upcoming Claude AI model will feature an embedded “imperceptible watermark” in the text it produces. This could complicate efforts to pass off AI-generated content as human-written.

As stated by Anthropic, these watermarks won’t alter the text’s meaning or readability. They will move along with the content when it’s copied and pasted, and some modifications may still show up in the results.

A watermark serves as a barely noticeable pattern or hidden marker within digital content, like a document or image, to help pinpoint the source and deter unauthorized replication.

This feature aligns with Anthropic’s commitment to enhancing transparency in accordance with the European Union AI Act. Future Claude models will incorporate watermarks from the outset, and the company is also looking to implement this capability in older versions. The markings will apply to Claude-generated material globally, including access through various cloud services. Additionally, Anthropic plans to offer tools to third parties for watermark detection.

The introduction of watermarks might transform the publishing landscape, notably in light of ongoing debates surrounding AI-generated texts. For instance, last month, a book distributor pulled support for a crime novel title, Please Call Me, I’ll Hide My Body, sparking discussions about whether AI had been used in its creation. Ultimately, fourteen publishers expressed interest, and a deal was finalized, following an apology from the distributor. Author Jerry Farad denied engaging AI in the novel’s writing.

Earlier this year, concerns arose with Hachette’s release of Mia Ballard’s horror novel Shy Girl, as some suspected the inclusion of AI-generated sentences. Ballard clarified that while she did not employ AI in her writing, a freelance editor may have introduced AI content without her awareness.

In another case, a book focused on truth in the AI era contained fabricated quotes. The author, Rosenbaum, acknowledged the presence of “inappropriate and composite quotations” and initiated an investigation into the errors. He described the inclusion of misquotes as coincidental, asserting no intent to misrepresent views during the writing process. He admitted to utilizing AI tools like ChatGPT and Claude and expressed accountability for the mistakes, working with his editors to rectify the issues for future editions.

Published by BenBella Books and distributed by Simon and Schuster, the book’s publishers have not provided comments on the situation.

With Claude’s new watermark, publishers, educational institutions, and others will have an additional resource for scrutinizing similar cases, making it more challenging to represent AI-generated writings as entirely human-created. However, the company acknowledged the system’s limitations, noting that if Claude’s output undergoes significant edits or is combined with other writings, the watermark might become undetectable. Furthermore, finding a watermark does not confirm Claude’s authorship, as even mild AI usage, like proofreading, can leave traces.

Anthropic is the second prominent AI organization to adopt watermarking in text; earlier this year, Google DeepMind announced the intention to implement SynthID technology for watermarking text and videos created through its Gemini platform. DeepMind had already introduced a image-oriented version of this tool in 2023.

The ever-changing relationship between work and AI continues to draw attention.

Facebook
Twitter
LinkedIn
Reddit
Telegram
WhatsApp

Related News