구글은 검색 센터 문서와 블로그에서 AI로 쓴 글에 대해 여러 번 답했습니다. 요지는 AI를 썼는지보다 누구를 위해 썼는지를 본다는 것입니다. 다만 이 글은 구글 검색 기준입니다. 네이버 검색의 기준은 이 문서들에 나오지 않습니다.

네이버 지식iN에서 'AI 블로그'로 검색되는 질문이 9월 24일 기준 5만5천 건이 넘습니다. 'AI 글쓰기'도 4,788건입니다.

■ AI를 썼다는 것만으로는 문제가 아니다

구글은 2023년 2월 검색 센터 블로그의 문답에서 AI 또는 자동화의 올바른 사용은 가이드라인에 위배되지 않는다고 밝혔습니다. 검색 순위 조작이 주목적인 콘텐츠 생성은 스팸 정책 위반이지만, 그렇지 않은 용도로 AI를 쓰는 것은 위반이 아니라는 설명입니다.

반대로 AI를 썼다고 득을 보지도 않습니다. 같은 문답에서 구글은 AI를 사용한다고 해서 콘텐츠에 특별한 이점이 있는 것은 아니며, 다른 콘텐츠와 같은 콘텐츠일 뿐이라고 답했습니다.

AI를 써야 하느냐는 질문에는 이렇게 답했습니다. 유용하고 독창적인 콘텐츠를 만드는 데 AI가 필요하다면 써도 좋지만, 저렴한 비용으로 손쉽게 검색 순위를 조작할 수 있다고 생각한다면 쓰지 말라는 것입니다.

■ 선을 넘는 지점은 '대량'과 '가치 없음'

그럼 어디서부터 문제가 될까요?

구글의 생성형 AI 콘텐츠 안내 문서는 생성형 AI가 주제를 조사하거나 원본 콘텐츠에 구조를 더할 때 특히 유용하다고 적었습니다. 하지만 사용자에 대한 가치 창출 없이 많은 페이지를 만들면 '확장된 콘텐츠 악용'에 대한 스팸 정책을 위반할 수 있다고 경고합니다.

스팸 정책 문서가 든 예는 이렇습니다. 다른 콘텐츠를 스크래핑해 동의어 분석이나 번역 같은 자동 변환으로 가치가 거의 없는 페이지를 여럿 만드는 것, 가치 창출 없이 여러 웹페이지의 내용을 병합하거나 결합하는 것, 검색 키워드가 들어 있어도 독자에게는 거의 의미가 없는 페이지를 여럿 만드는 것입니다.

이 정의에는 '생성 방법에 관계없이'라는 표현이 들어 있습니다. 사람이 썼든 AI가 썼든, 가치 없는 글을 대량으로 찍어 내는 것 자체를 문제로 본다는 뜻입니다.

■ 스스로 점검할 질문

구글의 '유용하고 신뢰할 수 있는 사용자 중심 콘텐츠 만들기' 문서에는 '예'라고 답하면 위험 신호로 받아들이라는 질문 목록이 있습니다. AI로 블로그를 운영할 때 걸리기 쉬운 것을 추리면 이렇습니다.

  • 광범위한 자동화를 사용해 다양한 주제에 대한 콘텐츠를 제작하고 있나요?
  • 다른 사람의 이야기를 주로 요약하고 있으며 가치를 크게 더하지 않나요?
  • 인기 있는 주제에 관해서만 쓰고, 기존 독자를 위한 글은 쓰지 않나요?
  • 콘텐츠가 크게 바뀌지 않았는데도 최신 페이지처럼 보이도록 날짜를 바꾸나요?
  • 구글이 선호하는 단어 수가 있다고 듣고 분량을 맞추고 있나요?

마지막 질문에 대해 구글은 특별히 선호하는 단어 수가 없다고 적었습니다. 분량을 채우려고 AI에 글을 늘려 달라고 할 이유가 없다는 뜻입니다.

■ '누가, 어떻게, 왜'를 밝히라

같은 문서는 콘텐츠를 '누가, 어떻게, 왜'라는 측면에서 평가해 보라고 권합니다.

'누가'는 작성자입니다. 독자가 작성자 정보를 예상할 만한 글에는 정확한 바이라인을 표시하라고 권합니다. AI가 일부 관여했더라도 기자명 입력란에 AI를 작성자로 표시하지 말라는 문답도 있습니다.

'어떻게'는 제작 과정입니다. 자동화로 상당히 많은 콘텐츠를 만든다면 AI 생성 등 자동화를 쓴다는 사실이 방문자에게 드러나는지, 어떻게 썼는지 배경을 알려 주는지 스스로 물어보라고 합니다. 공개 문구는 독자가 '이 콘텐츠는 어떻게 만들었을까?'라고 생각할 만한 글에 쓰면 유용하다고 적었습니다.

구글이 가장 중요하다고 꼽은 것은 '왜'입니다. 사람들을 돕기 위해 만든 글이어야 하고, 주로 검색엔진 방문을 끌어오려고 만든 글은 구글 시스템이 보상하려는 콘텐츠가 아니라고 설명합니다.

제목과 설명문도 예외가 아닙니다. 구글은 콘텐츠를 자동으로 생성할 때 정확성, 품질, 관련성에 중점을 두라고 하면서, 여기에 검색 결과에 보일 수 있는 title 요소와 메타 설명, 구조화된 데이터, 이미지 대체 텍스트 같은 메타데이터도 포함된다고 적었습니다.

■ 정리하면

  • AI를 썼다는 것만으로는 구글 가이드라인 위반이 아닙니다 — 특별한 이점도 없습니다
  • 가치 없이 대량으로 만든 페이지는 스팸 정책 위반이 될 수 있습니다 — 만든 방법과 관계없습니다
  • 요약만 하고 더한 게 없거나, 날짜만 바꿔 새 글처럼 보이게 하는 건 위험 신호입니다
  • 구글이 선호하는 단어 수는 없습니다 — 분량을 채우려고 늘릴 필요가 없습니다
  • 작성자는 AI로 표시하지 않고, AI를 어떻게 썼는지는 독자가 궁금해할 만하면 밝힙니다
  • 구글 검색 기준입니다 — 네이버는 따로 확인해야 합니다

Google has answered questions about AI-written content repeatedly in its Search Central documentation and blog. The gist: it looks less at whether you used AI than at whom you wrote for. Note that this covers Google Search. Naver Search's standards do not appear in these documents.

On Naver's Q&A service, a search for 'AI blog' returns more than 55,000 questions as of September 24, and 'AI writing' 4,788.

■ Using AI is not a problem in itself

In a February 2023 Search Central blog Q&A, Google said appropriate use of AI or automation is not against its guidelines. Generating content primarily to manipulate search rankings violates its spam policies, but using AI for other purposes does not.

Nor does AI earn you anything. In the same Q&A, Google said using AI does not give content any special gains; it is just content like any other.

Asked whether one should use AI, Google said: if you see AI as essential to producing helpful, original content, go ahead; if you see it as an inexpensive, easy way to game search rankings, don't.

■ The line: volume without value

So where does it become a problem?

Google's guidance on generative AI content says it can be particularly useful for researching a topic or adding structure to original content. But generating many pages without adding value for users may violate its spam policy on scaled content abuse.

The spam policy's examples include scraping content and using automated transformations such as synonym swapping or translation to create many pages of little value; stitching or combining content from different pages without adding value; and creating many pages that contain search keywords but make little or no sense to a reader.

The definition includes the phrase 'no matter how it's created.' Human or AI, mass-producing content without value is the problem.

■ Questions to ask yourself

Google's page on creating helpful, reliable, people-first content lists questions that should be taken as a warning sign if you answer yes. The ones an AI-run blog is most likely to trip on:

  • Are you using extensive automation to produce content on many topics?
  • Are you mainly summarizing what others have to say without adding much value?
  • Are you writing about trending topics only, and not for your existing audience?
  • Are you changing dates to make pages seem fresh when the content has not substantially changed?
  • Are you writing to a word count because you heard Google has a preferred one?

On the last point, Google says it has no preferred word count - there is no reason to have AI pad a post to hit a length.

■ Show who, how and why

The same page recommends evaluating content in terms of who, how and why.

'Who' is the author. Where readers would expect it, Google recommends accurate bylines, and a Q&A says not to list AI as the author in a byline even if AI was involved.

'How' is the process. If you produce a substantial amount of content with automation, ask whether the use of automation, including AI generation, is self-evident to visitors and whether you explain how it was used. Disclosures are useful where readers might ask how the content was created.

Google calls 'why' the most important: content should exist to help people. Content made mainly to attract search engine visits is not what its systems seek to reward.

Titles and descriptions count too. Google says to focus on accuracy, quality and relevance when generating content automatically, including metadata that may appear in search results such as the title element, meta description, structured data and image alt text.

■ In short

  • Using AI alone does not break Google's guidelines - nor does it give special gains
  • Mass-producing pages without value may violate spam policy - however they were made
  • Summarizing without adding value, or changing dates to look fresh, are warning signs
  • Google has no preferred word count - no need to pad
  • Do not list AI as the author; disclose how AI was used where readers would wonder
  • This is Google's standard - check Naver separately