Every GEO check the scanner runs

The scanner runs 40 explicit checks across 12 weighted dimensions. Every one of them is listed here with the exact rule it applies, so a score can be argued with rather than taken on trust. Nothing on these pages is a summary of the method — it is the method.

A check is worth points inside its dimension, and each dimension carries a share of the total. A dimension is not a checklist of equal items: the weights are published below and the reasoning for each one is on its page.

AI Crawler Access

16% of total

If a crawler cannot read the page, nothing else on this list can matter. This is the only dimension where a failure makes the rest of the score moot.

Machine Readability

12% of total

JSON-LD is how a model binds your brand name, domain and product into one entity instead of inferring three unrelated strings.

  • At least one <script type="application/ld+json"> block parses as JSON and declares at least one @-keyword (@context, @type, @graph or @id).

  • Parsed JSON-LD contains an @type of Organization, WebSite, Person or LocalBusiness.

  • Parsed JSON-LD contains an @type of FAQPage, Article, BlogPosting, HowTo, Product, SoftwareApplication or BreadcrumbList.

Content Depth

11% of total

Generative engines select sources that answer a question thoroughly. Thin pages are rarely retrievable regardless of how well they are marked up.

Citability & Evidence

11% of total

Statistics, quotations and cited sources are the interventions with the largest measured effect in the published GEO research.

  • Body text contains at least 5 numeric claims: percentages, currency amounts, multipliers of the form 3.2x, thousands-separated figures, or raw numbers of five digits or more.

  • The page contains at least one <blockquote> element.

  • The page links to at least 2 external hosts on the authority list: arxiv.org, doi.org, nature.com, science.org, acm.org, ieee.org, springer.com, sciencedirect.com, any .gov or .edu host, wikipedia.org, github.com, developer.mozilla.org, ...

  • The page declares a rel=canonical link.

Answer Readiness

10% of total

Answers are extracted as spans, not pages. A question-shaped heading followed by an immediate answer is the easiest thing for a retrieval system to lift intact.

Trust & Authority

10% of total

E-E-A-T signals decide whether a model treats a claim as safe to repeat rather than something it should hedge.

  • The homepage was retrieved over https://, rather than only over plain http://.

  • Hrefs on the page include both an about-style path (about, company, team, who-we-are) and a contact-style path (contact, support) or a mailto: link.

  • Authorship is signalled by any of: a Person node in parsed JSON-LD, an author property with a non-empty value in parsed JSON-LD, rel="author", or a visible byline matching "by Firstname Lastname".

  • Parsed JSON-LD contains a sameAs property with a non-empty value.

Semantic Structure

8% of total

Heading hierarchy and landmark elements are how a parser locates section boundaries at all.

Metadata & Discoverability

7% of total

Title, description and canonical control what a search or answer surface can show about the page.

AI Context Files

5% of total

A cheap and optional signal. Google has stated it does not use llms.txt in Search, and crawler support is inconsistent.

Freshness

5% of total

Dated content is deprioritised in generated answers, and an undated page gives an engine nothing to reason about.

International Readiness

3% of total

Language and region markup decides which language market a page can be retrieved in at all.

Delivery & Mobile

2% of total

A slow or non-mobile-readable page is dropped before any content analysis happens.

Questions about the checks

Can I see the rule without running a scan?

Yes, and that is the point of this section. Every rule, its point value and the exact condition that makes it pass are published, so you can read the method before deciding whether the score is worth anything.

Are all checks applied to every page?

No. A small number are marked not applicable to particular kinds of page, such as a policy or contact page, and are removed from the calculation rather than counted as failures. Each exclusion is listed on the page for the check it applies to.

What happens when a check cannot be evaluated?

It says so rather than guessing. A refusal aimed at our scanner is reported as unverified, not as a failure of your site, because those are different findings and only one of them is yours to fix.

Run a free GEO audit on your own site →

The dimension weights, the A–F bands and the parts of the picture a single-URL scan cannot see are all on the methodology page.