[レベル: 初級]
Google は、検索セントラルサイトで公開している次の 2 つのドキュメントを更新しました。
ドキュメント更新内容
各ドキュメントの更新内容を解説します。
なお、この記事を書いている時点では日本語ドキュメントは未更新です。
したがって更新内容は、オリジナルの英語ドキュメントに基づきます。
ウェブサイトで生成 AI によるコンテンツを使用するための Google 検索のガイダンス
「ウェブサイトで生成 AI によるコンテンツを使用するための Google 検索のガイダンス」ドキュメントの「正確性、品質、関連性を優先する (Focus on accuracy, quality, and relevance) 」セクションに注釈の記述が追加されました。
以前の記述:
When creating content for the web, focus on accuracy, quality, and relevance, especially when automatically generating the content. This includes metadata like
<title>elements, meta description elements, structured data, and alternate texts for images, which can appear in Search results.
更新後の記述:
When creating content for the web, focus on accuracy, quality, and relevance, especially when automatically generating the content. Keep in mind that generative models don’t retrieve facts, but predict a likely sequence of words based on their training data. Because of this, generative AI outputs may contain inaccuracies (also known as hallucinations). It is critical to manually factcheck and review all AI-generated content for accuracy and trustworthiness before publishing. This review also applies to metadata like
<title>elements, meta description elements, structured data, and alternate texts for images, which can appear in Search results.
ハイライトした箇所が追加部分です。
日本語に訳します。
生成モデルは事実を取得するのではなく、学習データに基づいて、もっともらしい単語の並びを予測するものであることに留意してください。そのため、生成 AI の出力には、不正確な情報(いわゆる「ハルシネーション」)が含まれる可能性があります。公開する前に、AI が生成したすべてのコンテンツについて、人の手でファクトチェックとレビューを行い、正確性と信頼性を確認することが不可欠です。
基本的に、LLM は、学習データに基づく予測モデルにすぎず事実を出力しているわけではありません。
ハルシネーションが常に付きまとうので、目視でのチェックが必須だと注意喚起しています。
⚠️ハルシネーションを抑制するための技術が Grounding(検索の場合は、主として RAG)だけど、ここでは原則論を述べている
公式の Search Central Live や一般団体主催のカンファレンスなどのプレゼンテーションでは、たいていこの注意点に触れています。
そこで同期を目的に、ドキュメントを更新しました。
有用で信頼性の高い、ユーザーを第一に考えたコンテンツの作成
「有用で信頼性の高い、ユーザーを第一に考えたコンテンツの作成」ドキュメントには、複数の記述およびセクションが追加されました。
「E-E-A-T と品質評価ガイドラインについて (Get to know E-E-A-T and the quality rater guidelines)」セクションに次の段落が加わりました。
Beyond E-E-A-T, the Search Quality Evaluator Guidelines emphasize that one of the most critical factors for assessing page quality is the quality of the main content, specifically how effectively it allows the page to achieve its purpose and provide a satisfying user experience.
日本語訳です。
E-E-A-T に加えて、検索品質評価ガイドラインでは、ページ品質を評価するうえで最も重要な要素の 1 つとして、メインコンテンツの品質が重視されています。具体的には、そのコンテンツがページの目的達成にどれだけ効果的に貢献し、ユーザーに満足のいく体験を提供できているかが評価されます。
メインコンテンツの品質に関する記述です。
これを補足する形で、「Main content(メインコンテンツ)」セクションが挿入されました。
Main content
Google Search defines “main content” as any part of the webpage that directly helps the page achieve its purpose. Depending on the primary intent of the page, the main content directly fulfills the page’s purpose and could include:
- Primary text and media: The main text, articles, images, audio, or video files on the page.
- Interactive features: Built-in functionality such as calculators, online tools, games, or search functionality.
- User-generated contributions: Forum discussions, customer reviews, comments, or user-uploaded media when they directly fulfill the page’s purpose (for example, when the page is about discussing a particular user question).
- Tabbed or expanded sections: Information housed behind interactive tabs (such as product specifications, safety notes, or user reviews).
- Page titles and headings: The visible title and headings throughout a page that summarize its topic and helps users make informed decisions about visiting.
When reviewing content quality, Search Quality raters are trained to evaluate these fundamental attributes of the main content:
- Effort: The extent to which human work went into creating the content or the systems powering it. For example, writing original analysis, translating a poem manually, or building a custom interactive map for a game represents high effort. Conversely, automatically generating pages from feeds or using generative AI to produce large amounts of text without manual oversight or curation represents little to no effort. Attribution or giving credit to other sources doesn’t replace the need for original effort.
- Originality: The extent to which the content offers unique, original information or perspectives that aren’t already available on other websites.
- Talent or skill: Whether the content shows the talent, skill, or expertise necessary to provide a satisfying experience for visitors (such as clear writing, well-produced video, or functional page tools). Keep in mind that not all content requires specialized expertise. For example, a person sharing their experience about finding a new way to clear their driveway of snow generally doesn’t necessarily require specialized expertise.
- Accuracy: For informational pages, the content should be factually accurate. For topics that could significantly impact people’s lives or well-being (YMYL topics), the content must be highly accurate and consistent with established expert consensus.
日本語に訳します。
メインコンテンツ
Google 検索では、「メインコンテンツ」を、ページの目的達成に直接役立つウェブページ上のあらゆる部分と定義しています。ページの主な意図に応じて、メインコンテンツはページの目的を直接満たすものであり、次のようなものが含まれます。
- 主要なテキストやメディア: ページ上の本文、記事、画像、音声、動画ファイル
- インタラクティブな機能: 計算ツール、オンラインツール、ゲーム、検索機能など、ページに組み込まれた機能
- ユーザー生成コンテンツ: フォーラムでの議論、顧客レビュー、コメント、ユーザーがアップロードしたメディアなど、それらがページの目的を直接満たす場合。たとえば、特定のユーザーの質問について議論することが目的のページなどが該当します
- タブ内や展開式セクションの情報: 商品仕様、安全上の注意、ユーザーレビューなど、インタラクティブなタブの背後に格納されている情報
- ページタイトルと見出し: ページのトピックを要約し、ユーザーがそのページを訪問するかどうかを判断するのに役立つ、ページ内の目に見えるタイトルや見出し
コンテンツ品質を評価する際、次の基本的な属性をメインコンテンツについて評価するように検索品質評価者は訓練されています。
- 労力: コンテンツや、それを支えるシステムの作成に、どの程度人間の作業が費やされているか。たとえば、独自の分析を書くこと、詩を人の手で翻訳すること、ゲーム向けに独自のインタラクティブマップを作成することは、高い労力を示します。一方、フィードからページを自動生成したり、人による監督やキュレーションなしに生成 AI を使って大量の文章を作成したりすることは、ほとんど、あるいはまったく労力がかけられていないとみなされます。他の情報源を明記したりクレジットを付与したりしても、独自の労力が必要であることに変わりはありません
- 独自性: 他のウェブサイトですでに得られるものではない、独自の情報や視点をどの程度提供しているか
- 才能またはスキル: 訪問者に満足のいく体験を提供するために必要な才能、スキル、専門性がコンテンツに表れているか。たとえば、明快な文章、質の高い動画、正常に機能するページ内ツールなどが該当します。ただし、すべてのコンテンツに専門的な知識が必要なわけではありません。たとえば、自宅の車道の雪を除去する新しい方法を見つけた経験を共有する場合、必ずしも専門的な知識は必要ありません
- 正確性: 情報提供を目的とするページでは、コンテンツは事実として正確である必要があります。人々の生活や健康に重大な影響を与える可能性のあるトピック(YMYL トピック)では、コンテンツには非常に高い正確性が求められ、確立された専門家の見解と一致している必要があります
新しい情報ではなく、品質評価ガイドラインからの抜粋です。
「有用で信頼性の高い、ユーザーを第一に考えたコンテンツの作成」ドキュメントには、さらに更新が入っています。
「コンテンツに関する「誰が、どのように、なぜ」を考える (Ask “Who, How, and Why” about your content)」セクションに次の段落が追加されました。
However, avoid using deceptive authorship information. Fabricating creator profiles (such as by using AI-generated headshots, made-up names, or false credentials to make content appear as if it was written by human experts) is a form of deception. Any form of deception makes a page untrustworthy to both users and our automated quality systems, and is a signal of a low-quality page.
ただし、作成者情報を偽ることは避けてください。たとえば、AI で生成した顔写真、架空の名前、虚偽の経歴を使って、あたかも人間の専門家が執筆したコンテンツであるかのように見せるために、架空の作成者プロフィールを用意することは、欺瞞行為にあたります。どのような形であれ、欺瞞行為があると、そのページはユーザーにとっても Google の自動品質評価システムにとっても信頼できないものとなり、低品質なページであることを示すシグナルとなります。
架空の専門家をでっちあげることで、E-E-A-T を偽ってコンテンツを公開する施策を披露したセッションがどこかのカンファレンスであったことを思い出しました。
検索スパムにとどまらず、純粋に詐欺行為ですよね、これ。
◇ ◇ ◇
2 つのドキュメントの更新は、SEO をきちんと理解している人にとっては目新しい内容ではありません。
それでも、より広く周知させるために公式ドキュメントとして公開したのではないかと思われます。
