Claude’s new Scarlet Letter watermark is invisible—for now



“We’re adding marking to Claude’s output to comply with the EU AI Act, and other labs are taking similar steps,” Anthropic said. “It’s hard to identify AI-generated text, and this gives people better tools for identification. Text from supported Claude models, including output from Claude Code, will carry an invisible watermark, and it doesn’t change the meaning, quality, or readability of Claude’s responses. We also plan to ship a text detection API so users can do more of this themselves.”

The fun is just starting

In the EU, transparency requirements are meant to ensure AI tools like Claude don’t upset “the integrity and trust in the information ecosystem, raising new risks of misinformation and manipulation at scale, fraud, impersonation, and consumer deception.” One EU support article forecasted that the obligations would be the “primary compliance challenge” for many AI firms.

“People should know when they are interacting with AI or exposed to AI-generated content,” the European Commission’s guidelines said. “This will help them make informed decisions, calibrate their trust and reliance on AI, and avoid misinformation or deception.” Still, it’s hard to square this with the fact that a wholly generated article on a matter of public interest gets a watermark, but not a reader-facing label, if an editor properly reviews it.

AI firms like Anthropic are best positioned to develop watermarking solutions, the EU expects, since AI moves fast and there will be an ongoing “need for new methods and techniques to trace origin of information.”

But that largely leaves the societal value of such marks up to tech firms to decide, with the EU only stipulating that “techniques and methods should be sufficiently reliable, interoperable, effective and robust as far as this is technically feasible.”

In its post, Anthropic said it plans to continue working on its watermarks and detection methods that meet the EU’s demands. If Claude’s labels fail, the AI Act carries steep penalties for violations, including fines up to 15 million euros, or 3 percent of a company’s worldwide annual revenue.

Ars Editor-in-Chief Ken Fisher contributed to this report. This story was updated on August 13 to add a statement from Anthropic.



Source link

  • Related Posts

    MCP for agent-to-agent comms may be the riskiest protocol you’ve never heard of

    “AI agents give attackers a fresh set of connections to walk across,” Douglas McKee, director of vulnerability intelligence at Rapid7, told Ars. “Someone plants text in content, an agent will…

    Continue reading
    Lucid Motors’ EV output falls to lowest level in almost two years

    Lucid Motors built 2,954 electric vehicles (EVs) in the third quarter of this year, a 54% drop from a year ago, as the company purposely limits production to better meet…

    Continue reading

    Leave a Reply

    Your email address will not be published. Required fields are marked *

    You Missed

    Brandon Sanderson Books Are Buy 2, Get 1 Free Today

    Brandon Sanderson Books Are Buy 2, Get 1 Free Today

    Supreme Court wrestles with energy companies’ bid to block major climate-change lawsuit

    Supreme Court wrestles with energy companies’ bid to block major climate-change lawsuit

    MCP for agent-to-agent comms may be the riskiest protocol you’ve never heard of

    MCP for agent-to-agent comms may be the riskiest protocol you’ve never heard of

    Live From Paris: 4 Fashion Week Shows Editors Can’t Stop Talking About

    Live From Paris: 4 Fashion Week Shows Editors Can’t Stop Talking About

    SmartCentres Real Estate Investment Trust to Release 2026 Third Quarter Results and Host Conference Call

    In Photos: British Airways Unveils All-New Airbus A380 Interior With World’s Largest Business Class Cabin

    In Photos: British Airways Unveils All-New Airbus A380 Interior With World’s Largest Business Class Cabin