Scorecard · 20 criteria · 6 pages

AEO Content Scorecard

20-criterion scorecard for evaluating whether a piece of content is built to be cited by ChatGPT, Perplexity, Google AI Overviews, Gemini, and Claude. Print, score, fix.

Author: Martin Vassilev · Updated 2026-04-25 · Reviewed by the Toronto SEO Editorial Board

Score a page automatically

How to use this scorecard

  1. Pick the page you want to evaluate.
  2. Walk through each criterion, scoring 1 (yes), 0.5 (partial), or 0 (no).
  3. Multiply each score by its weight.
  4. Sum the weighted scores. Total possible: 74 points.
  5. Above 59: production-ready for AI citation. Below 44: structural rebuild needed before AEO will move.

Scoring system: this is a deliberately weighted heuristic. The weights are based on observation across our portfolio — not a published Google ranking factor. Treat the absolute number as a comparison tool across your own pages, not as a regulator-grade benchmark.

Structure

6 criteria · 22 pts
  1. 1

    Does the content open with a self-contained 40–80 word answer to the primary query?

    5 pts
    YesPartialNo
    Yes means

    AI engines can extract the lead as a citation snippet without context-hunting deeper into the page.

    No means

    The page may still be indexed, but the citation will likely go to a competitor whose lead paragraph is more extractable.

  2. 2

    Is the H1 a verbatim or near-verbatim match to the user query intent?

    4 pts
    YesPartialNo
    Yes means

    The H1 acts as a strong relevance signal for both Google and AI engines doing passage extraction.

    No means

    The page is competing with itself — Google has to guess which query the page actually serves.

  3. 3

    Are H2s phrased as natural-language sub-questions or scannable nouns?

    4 pts
    YesPartialNo
    Yes means

    AI engines extract H2-level headings as candidate facets to cite or quote.

    No means

    Decorative or marketing-style H2s ('Discover the magic of...') signal low information density.

  4. 4

    Are key facts and figures formatted as scannable lists, tables, or callouts?

    3 pts
    YesPartialNo
    Yes means

    Structured data extracts cleanly into AI engine summaries and Google AIO snippets.

    No means

    Buried-in-prose stats are systematically harder to cite than the same stats in a bulleted list or table.

  5. 5

    Does the page contain a genuine FAQ section with 3+ user-asked questions?

    3 pts
    YesPartialNo
    Yes means

    FAQPage schema becomes legitimately applicable; AI engines preferentially cite FAQ-marked answers.

    No means

    If you fabricate FAQs to qualify for FAQPage schema, that's structured-data spam — Google demotes it.

  6. 6

    Is the visible content depth proportional to the SERP norm (no 400-word page on a 2,000-word query)?

    3 pts
    YesPartialNo
    Yes means

    The page can plausibly displace the existing top-3 in the SERP it's competing in.

    No means

    Underbuilt content cannot displace overbuilt content, regardless of how clean the technical SEO is.

Authority

5 criteria · 20 pts
  1. 1

    Does the page have a named, real, credentialed author with a linked bio?

    5 pts
    YesPartialNo
    Yes means

    E-E-A-T signal is present and consistent. AI engines disproportionately cite authored content.

    No means

    Anonymous content is increasingly demoted in AI citation ranking. 'By the Editor' is a red flag.

  2. 2

    Is the author tied to a stable Person @id used across every other article they write?

    4 pts
    YesPartialNo
    Yes means

    The Person entity consolidates across the site, building cumulative E-E-A-T weight.

    No means

    Fragmented author entities dilute E-E-A-T across multiple synthetic authors.

  3. 3

    Is there a named reviewer or fact-checker, also tied to a stable Person @id?

    4 pts
    YesPartialNo
    Yes means

    Fact-checking signal is present in schema and visible on-page; YMYL content particularly benefits.

    No means

    On YMYL queries (medical, legal, financial), the citation gap to fact-checked content is widening.

  4. 4

    Does the content cite primary sources (research, regulators, original data) with hyperlinks?

    4 pts
    YesPartialNo
    Yes means

    AI engines preferentially cite content that itself cites verifiable primary sources.

    No means

    Uncited claims look like AI-generated filler — and are increasingly demoted in citation rankings.

  5. 5

    Does the content include first-party data, observation, or original analysis?

    3 pts
    YesPartialNo
    Yes means

    Information gain is high; AI engines prefer to cite content that adds something not findable elsewhere.

    No means

    Pure synthesis of competitor content is structurally lower in citation rate than original-data content.

Citation Surface

5 criteria · 17 pts
  1. 1

    Does the page have appropriate Article + Author + Publisher schema with stable @ids?

    4 pts
    YesPartialNo
    Yes means

    Structured data is consolidated; AI engines can resolve the entity graph cleanly.

    No means

    Schema fragmentation or absence caps citation eligibility regardless of content quality.

  2. 2

    Does the page use FAQPage schema where genuinely applicable?

    4 pts
    YesPartialNo
    Yes means

    Question/answer pairs are individually citation-eligible by AI engines.

    No means

    Eligible FAQ content is going un-cited.

  3. 3

    Does the page use HowTo schema if it's genuinely procedural?

    3 pts
    YesPartialNo
    Yes means

    Step-by-step content is eligible for procedural rich results and AI procedure citations.

    No means

    Either eligible HowTo content is missing markup, or non-procedural content is being marked up (don't do that).

  4. 4

    Are images alt-tagged with descriptive context, not just keywords?

    3 pts
    YesPartialNo
    Yes means

    Image-search and AI-engine image understanding can use the page as a citation surface.

    No means

    Image-driven citation surfaces (Google Lens, AI image queries) are leaving signal on the table.

  5. 5

    Are external citations (where you reference others) made via descriptive anchor text and proper attribution?

    3 pts
    YesPartialNo
    Yes means

    AI engines reward reciprocal citation behaviour with higher in-citation rates.

    No means

    Citation manipulation patterns ('source: Wikipedia' with no link) are increasingly detected.

Crawlability

4 criteria · 15 pts
  1. 1

    Is robots.txt explicitly allowing the AI crawlers you want to be cited by?

    5 pts
    YesPartialNo
    Yes means

    ChatGPT-User, OAI-SearchBot, PerplexityBot, ClaudeBot, GoogleOther are not blocked at the path level.

    No means

    Default Cloudflare and CMS hardening configs frequently block AI crawlers — you may be invisible.

  2. 2

    Does the page render full content in the initial HTML (not just JS-rendered)?

    4 pts
    YesPartialNo
    Yes means

    AI crawlers (which often don't execute JavaScript) can read your content.

    No means

    JS-rendered content is systematically under-cited by AI engines that don't render JS reliably.

  3. 3

    Is page response time under 3 seconds to first byte?

    3 pts
    YesPartialNo
    Yes means

    AI crawlers (which have shorter timeouts than Googlebot) can complete the fetch.

    No means

    Slow pages get systematically skipped from AI engine indexes.

  4. 4

    Is the page reachable from the homepage in <3 clicks via internal links?

    3 pts
    YesPartialNo
    Yes means

    Crawl frequency and citation eligibility both correlate with internal link depth.

    No means

    Orphaned or deeply-buried pages cite at a fraction of the rate of well-linked pages of comparable quality.