사람이 쓴 글을 클로드로 다듬기만 해도 표식이 남을 수 있다.

앤스로픽이 클로드가 만든 모든 글에 사람 눈에 보이지 않는 워터마크를 심기로 했다는 보도가 나왔다.

무엇이 어디에 심기나

워터마크는 텍스트의 의미나 가독성을 바꾸지 않으면서 문장 안에 직접 심긴다. 복사해 붙여넣거나 일부를 고쳐도 표식이 일부 남도록 설계됐다.

적용 대상은 8월 2일 이후 출시된 모든 모델이다. 이전 모델에도 순차적으로 적용할 예정이라고 한다.

유럽연합 AI법의 투명성 의무에 따른 조치이고, 전 세계 이용자에게 똑같이 적용된다. 앤스로픽은 제3자가 워터마크를 탐지할 수 있는 도구도 제공할 계획이다.

텍스트에 워터마크를 넣은 주요 AI 기업으로는 두 번째다. 구글 딥마인드가 2024년부터 신스ID로 제미나이가 만든 텍스트와 영상에 표식을 넣어왔고, 이미지는 2023년부터였다.

출판·교육이 먼저 반응할 이유

AI 집필 의혹으로 계약이 뒤집히는 일이 실제로 있었다.

지난달 범죄소설 '전화해, 내가 시체는 내가 치울게'는 14개 출판사가 입찰 경쟁을 벌인 끝에 계약됐지만, 저자 제리 팔라데의 AI 집필 의혹이 불거지며 계약이 철회됐다.

미국 출판사 아셰트도 미아 밸러드의 공포소설 '수줍은 소녀' 출간을 중단했다. 밸러드는 뉴욕타임스에 프리랜서 편집자가 자신도 모르게 AI 생성 내용을 끼워 넣었다고 해명했다.

과제물이나 원고에 표식이 남아 있는지 확인할 수단이 생기는 셈이다.

탐지됐다고 AI가 쓴 건 아니다

사람이 쓴 글을 클로드로 교정하거나 번역만 해도 워터마크가 남을 수 있다. 그래서 표식이 나왔다는 사실만으로 글 전체를 AI가 썼다고 단정할 수 없다.

반대쪽 구멍도 있다. 앤스로픽은 워터마크가 들어간 글이라도 대대적인 편집이나 의역, 번역을 거치거나 다른 글과 섞이면 탐지가 어려워질 수 있다고 밝혔다.

정리하면 이렇다. 있는데 못 잡을 수도 있고, 잡혔는데 사람이 쓴 글일 수도 있다.

판별 도구가 쉽지 않다는 건 전례가 있다. 오픈AI는 2023년 자체 개발한 AI 텍스트 판별기의 정확도가 낮다는 이유로 서비스를 접었다.

출처를 짚고 가자. 이 내용은 10일 비즈니스 인사이더 보도를 국내 언론이 전한 것이고, 앤스로픽 공식 뉴스 페이지에는 8월 11일 기준 이 발표가 올라와 있지 않다.

이전 모델에 언제부터 적용할지, 탐지 도구를 언제 공개할지도 아직 나오지 않았다.

표식이 늘어나는 방향은 정해진 듯하다. 다만 그 표식을 어떻게 읽어야 하는지는 아직 아무도 정하지 않았다..

Run a piece of your own writing through Claude just to clean it up, and a mark may stay in it.

Anthropic is reported to be embedding an invisible watermark in everything Claude writes.

What goes in, and where

The watermark is embedded directly inside sentences without changing their meaning or readability. It is designed so that part of the mark survives copy-paste and partial edits.

It applies to every model released after August 2, and Anthropic reportedly plans to roll it out to earlier models over time.

The move follows transparency obligations under the EU AI Act, and applies identically to users worldwide. Anthropic also plans to provide a tool that lets third parties detect the watermark.

This makes Anthropic the second major AI company to watermark text. Google DeepMind has used SynthID on Gemini's text and video since 2024, and on images since 2023.

Why publishing and education react first

Contracts have already been reversed over suspected AI authorship.

Last month the crime novel Call Me, I'll Hide the Body was signed after a bidding contest among 14 publishers, then had its contract withdrawn when suspicions arose that author Jerry Palade had written it with AI.

The US publisher Hachette also halted publication of Mia Ballard's horror novel Shy Girl. Ballard told the New York Times that a freelance editor had inserted AI-generated material without her knowledge.

A watermark gives schools and publishers a way to check whether a mark is present in a submitted paper or manuscript.

A detection is not proof of authorship

If you write something yourself and only use Claude to edit or translate it, the watermark can still be there. So finding a mark does not establish that AI wrote the piece.

The gap runs the other way too. Anthropic says detection can become difficult if watermarked text is heavily edited, paraphrased, translated, or mixed with other writing.

So it can be present and missed, or detected on writing a person actually wrote.

Detection tools have a track record here. OpenAI shut down its own AI text classifier in 2023, citing low accuracy.

A note on sourcing. This comes from a Business Insider report on the 10th, relayed by Korean media, and as of August 11 no such announcement appears on Anthropic's official news page.

When earlier models get the watermark, and when the detection tool ships, have not been stated either.

More marking looks inevitable. What nobody has settled is how those marks should be read..

Sources · The Dong-A Ilbo