Our methodology

The AEO checklist

Every check we run, and why each one changes whether an answer engine cites you. We detect what a page actually is — its shape and its subject — then run the checklists that apply. A cyclone report and a lipstick page are not graded the same way.

79checks
5page shapes
13verticals
Detection reads the page's schema.org markup first, then falls back to structure and vocabulary. Every audit reports which checklist it chose and the evidence behind that call, so you can disagree with it.

Every page

15 checks

Run against every page regardless of what it is.

Answer in the opening lines

Decisive

Models retrieve a passage, not a page. If the answer is not in the first two sentences of a section, it usually is not in the chunk that gets quoted.

Content present without JavaScript

Decisive

Googlebot renders JavaScript; GPTBot, ClaudeBot and PerplexityBot do not. JS-injected content is invisible to every answer engine at once.

Headings phrased as real questions

Important

Retrieval matches a user’s question against your headings. A heading that is already the question is the strongest match you can offer.

Claims attributed to named sources

Decisive

Controlled testing (the Princeton GEO study) found citing authoritative sources measurably increases how often a page is quoted. Models repeat attributed claims more readily than floating ones.

Specific numbers rather than adjectives

Important

A number is self-contained and survives being lifted out of context. "Traffic fell sharply" is unusable to an answer engine; "traffic fell 58%" is quotable on its own.

Entities named, not implied by pronouns

Important

A retrieved chunk arrives without the paragraph that defined "it". Sections that lean on pronouns become unusable once separated.

Published and updated dates in markup

Important

Answer engines prefer sources they can date, and heavily discount undated pages on anything time-sensitive.

Structured data describing the page

Decisive

Roughly two-thirds of the pages Google AI Mode cites, and about seven in ten of the pages ChatGPT cites, carry schema markup. It is the clearest signal in the whole checklist: it tells the engine what the page is instead of making it guess.

Sections short enough to be quoted whole

Important

Engines retrieve a passage of roughly 100–300 words, not the page. One long undivided block gets split mid-thought and usually discarded; a page of tight, self-contained sections gives the engine something it can lift intact.

Original data or first-hand evidence

Important

Pages carrying original figures — your own testing, survey or internal data — are measurably more likely to be quoted, because they are the only source for that number. Content that only restates what other sites already say gives an engine no reason to pick it over them.

One H1 and a clean heading order

Worth having

The heading tree is how a parser works out which text belongs to which topic. Several H1s, or an H2 that jumps straight to H4, produce sections attached to the wrong heading — so the right answer gets filed under the wrong question.

Images described in text

Worth having

The crawlers behind ChatGPT, Claude and Perplexity read HTML, not pictures. Anything that exists only inside an image — a spec table, a chart, a comparison — is simply absent unless the alt text says what it shows.

Page light enough for AI crawlers

Worth having

AI crawlers are far less patient than Googlebot — many give up after one to five seconds. A very heavy document risks being abandoned before the content is read, which costs the citation regardless of how good the writing is.

Not blocking snippets or AI use

Worth having

nosnippet and max-snippet:0 tell search engines not to show any text preview of this page at all — which also removes it as a candidate for AI Overviews and AI Mode. This is usually set by mistake, in a template, not per page.

Video content described in text

Worth having

AI crawlers do not watch video. Anything said only on camera — a demo, an explanation, a verdict — is invisible unless the surrounding text or a transcript restates it.

Article

2 checks

Added when the page is detected as this shape.

FAQ block

Important

FAQ markup maps one question directly onto one quotable answer — the single most extractable structure a page can offer.

Named author

Important

An identifiable author is a trust signal engines weigh when choosing between two otherwise-similar sources.

Product page

7 checks

Added when the page is detected as this shape.

Price readable in the HTML

Decisive

AI shopping answers quote price. If the price is injected by JavaScript, the engine reports the product without one — or skips it.

Product schema with offers

Decisive

Product markup removes guesswork: the engine reads price, currency and availability as declared values instead of inferring them from prose.

Review text actually in the HTML

Decisive

A declared count persuades humans; the review text persuades models. If you claim 1,200 reviews and render three, three is the entire evidence base an engine has.

Stock status in text

Important

Shopping answers filter on availability. An engine that cannot read stock status usually drops the product from the comparison.

Brand and model named explicitly

Decisive

Product questions are asked by brand and model. A page titled only "Wireless Earbuds" cannot be matched to the query.

Returns / warranty terms on the page

Worth having

"Can I return it" is among the most-asked pre-purchase questions, and it is usually buried on a separate policy page the engine never associates with the product.

Specifications as text

Important

Specs are what comparison questions are answered from. In an image or a script-built table, they do not exist to the engine.

Category / listing

2 checks

Added when the page is detected as this shape.

ItemList schema

Important

"Best X under Y" questions are answered from listing pages. ItemList tells the engine this is a ranked set rather than one long article.

Each entry named with its price

Decisive

A recommendation is only quotable if the engine can pair a specific product name with a specific price.

Review

4 checks

Added when the page is detected as this shape.

Verdict stated up front

Decisive

The question is "should I buy this". A verdict in the opening lines is the passage most likely to be lifted verbatim.

Pros and cons as text lists

Important

Pros/cons are pre-chunked, balanced, quotable statements — close to ideal retrieval units.

Testing methodology stated

Decisive

First-hand testing is what separates a review an engine will trust from marketing copy it will ignore.

Review schema with rating and author

Important

Review markup lets an engine attribute the verdict to a named reviewer and a numeric score rather than parsing it out of prose.

How-to / guide

3 checks

Added when the page is detected as this shape.

Steps as a real ordered list

Decisive

Step-by-step answers are reproduced almost verbatim by assistants — but only when the steps are list markup rather than paragraphs.

HowTo or Recipe schema

Important

This markup declares the steps, timings and materials explicitly, which is what voice and assistant answers read from.

Materials or prerequisites listed

Important

An assistant relaying instructions needs to state what is required before step one, or the answer is incomplete and gets passed over.

News

4 checks

Added when the page is detected as this subject.

Lead answers who / what / when / where

Decisive

News answers are assembled from the lead. A lead missing the when or where cannot be used to answer a factual question about the event.

NewsArticle schema with both timestamps

Decisive

On news queries, recency is close to decisive. Without machine-readable timestamps an engine cannot tell whether your version is the current one.

Primary source named

Decisive

Engines strongly prefer the outlet that names the official, agency or document over one that writes "sources said".

Dateline / location line

Worth having

A dateline is how a wire story declares where it was filed, which is exactly what location-scoped questions are matched against.

Health & Medical

5 checks

Added when the page is detected as this subject.

Author’s clinical credentials on the page

Decisive

Health is the category where engines are most conservative. An uncredentialed health page is routinely passed over in favour of one with a named clinician.

Medically reviewed, with a date

Decisive

A review line with a date is the clearest trust marker a health page can carry, and it is machine-readable.

Citations to primary literature

Decisive

Linking the study, WHO or CDC rather than another blog is what lets an engine treat a health claim as supported.

Medical disclaimer

Important

Its absence reads as a risk signal to an engine deciding whether to surface health guidance at all.

No absolute cure claims

Important

Absolute language ("cures", "guaranteed") is a suppression trigger for health content across every major engine.

Beauty & Personal Care

5 checks

Added when the page is detected as this subject.

Full ingredient list as text

Decisive

"Does it contain X", "is it safe for sensitive skin" are answered from the INCI list. As an image it does not exist to a crawler.

Skin/hair type suitability stated

Decisive

Nearly every beauty question is conditional — "for oily skin", "for curly hair". Without the qualifier the page cannot match the question.

How to use

Important

Application steps are a distinct, highly-asked question and a separate citation opportunity from the product description.

Claims substantiated

Important

"Dermatologically tested" with no testing body reads as marketing. Named substantiation is what makes a claim repeatable by an engine.

Shades / variants readable

Worth having

Shade availability is a common question, and it is usually rendered as swatch images with no text alternative.

E-commerce

1 check

Added when the page is detected as this subject.

Delivery and payment terms in text

Worth having

Shopping answers increasingly include delivery time and payment options; both are usually rendered by script or hidden behind a tab.

Entertainment

3 checks

Added when the page is detected as this subject.

Cast, director and title named

Decisive

Entertainment questions are almost entirely entity lookups. Unnamed people cannot be matched to "who directed…".

Where to watch and when it released

Decisive

"Where can I watch X" is the highest-volume question in this category and needs an explicit platform name and date.

Movie / TVSeries schema

Important

This markup declares the title, cast and release date as data rather than leaving them to be parsed out of prose.

Lifestyle & How-to

2 checks

Added when the page is detected as this subject.

Steps are self-contained

Important

Assistants relay one step at a time. A step reading "repeat with the rest" is meaningless once separated from its neighbours.

Time and difficulty stated

Important

"How long does it take" is asked of nearly every how-to, and is answerable only if you state it.

Product Review

3 checks

Added when the page is detected as this subject.

Compared against named alternatives

Decisive

"X vs Y" is how buying questions are actually asked. A review naming no alternative cannot be retrieved for any comparison.

Price stated, with the date it applied

Important

A price with no date goes stale silently, and an engine repeating a stale price is a reason to stop trusting the source.

Reviewer identified with relevant experience

Important

First-hand expertise is the differentiator engines use to separate a real review from aggregated marketing copy.

Fintech & Finance

4 checks

Added when the page is detected as this subject.

Fees, rates and charges stated as text

Decisive

"How much does X cost" and "what is the interest rate" are the single most-asked fintech questions, and the numbers are usually rendered by a calculator widget rather than existing as text.

Eligibility criteria stated

Important

"Who can apply" or "am I eligible" gates every financial-product decision. Without an explicit answer, an engine cannot tell a user whether the product even applies to them.

Regulatory / licensing disclosure

Decisive

Financial claims carry more scrutiny than most content. A named regulator or license number is what separates a credible source from marketing copy, for both readers and answer engines.

Risks stated, not just benefits

Important

A page that only lists upside reads as an advertisement. Named, specific risk disclosure is a trust signal and is often a legal requirement.

Real Estate

4 checks

Added when the page is detected as this subject.

Price and price-per-sq-ft stated

Decisive

"What does it cost" and "what is the rate per square foot" are the two numbers every property question needs, and listings often bury them in a downloadable brochure.

Location, configuration and possession stated

Decisive

"Is this in X locality", "what configuration is available" and "when can I move in" are asked of every property page. Vague location or a missing possession date makes the listing unusable for these queries.

RERA / regulatory registration where applicable

Important

In markets with a property regulator, a listed registration number is the clearest trust signal a listing can carry — and its absence is itself a question buyers ask.

Nearby amenities named

Important

"What is nearby" — schools, hospitals, transit — is a standard property question, and is only answerable if named rather than shown on an interactive map widget alone.

Automotive

4 checks

Added when the page is detected as this subject.

Core specifications as text

Decisive

"What engine does it have", "what is the mileage" — these are answered from a spec sheet. If it renders as an image or a JS-built table, the numbers do not exist to a crawler.

Variant and on-road price stated

Important

"Which variant is this" and "what does the on-road price come to" cannot be answered from an ex-showroom figure alone — a very common gap.

Safety rating or features named

Important

"Is it safe", "how many airbags" are pre-purchase questions that specifically need a named rating or feature list — general "safety" copy does not answer them.

Compared against named alternatives

Important

Automotive buying decisions are almost always comparative. A page naming no competing model cannot be retrieved for any "X vs Y" query.

Education & EdTech

4 checks

Added when the page is detected as this subject.

Eligibility, fees and duration stated

Decisive

"Can I apply", "what does it cost" and "how long does it take" are asked before anything else about a course, and are frequently split across a brochure PDF instead of the page itself.

Faculty or instructor credentials named

Important

Course quality is judged by who teaches it. An unnamed "expert faculty" claim carries no evidence an engine can repeat.

Learning outcomes or placement data stated

Important

"What will I be able to do after this" and "what are the outcomes" are the actual purchase questions — a syllabus list alone does not answer them.

Course schema present

Worth having

Course markup lets an engine read subject, level, instructor and duration as structured data instead of parsing marketing copy.

SaaS & Technology

4 checks

Added when the page is detected as this subject.

What it does and who it is for, stated plainly

Decisive

"What does this tool do" and "who is it for" are the first two questions any comparison or recommendation prompt needs answered — and are often buried under a hero tagline that says neither.

Pricing stated as text

Important

"How much does it cost" is asked constantly in tool-comparison prompts, and pricing pages frequently render the actual numbers via JavaScript widgets that a crawler never sees.

Alternatives or competitors named

Important

"X vs Y" and "alternatives to X" are extremely common SaaS-discovery prompts. A product page that never names a competitor cannot be surfaced for either.

Security / compliance stated where relevant

Worth having

B2B buying prompts frequently ask about compliance (SOC 2, GDPR). Its absence from the page is itself an answer an engine will report.

Music & Audio

3 checks

Added when the page is detected as this subject.

Artist, album and release date named

Decisive

"Who sings this", "what album is it from" and "when did it release" are the core entity questions for any music page, and are the first things an engine needs to ground an answer.

Official streaming links named

Important

"Where can I listen to this" is a direct-action question. Naming the platforms (Spotify, Apple Music) is what makes the page useful to point to, rather than just descriptive.

Music schema present

Worth having

MusicRecording/MusicAlbum markup declares artist, album and duration as structured data an engine can read directly.

See where a page of yours stands

Run one against this checklist. The first check is on us.

Run a check