Illucrum
Get in touch

By Szymon Kokot Published Updated 28 min read

GEO Checklist (2026): 51 Checks for AI Search Visibility

A GEO checklist tests whether AI systems can reach, read, trust and cite a business. GEO stands for generative engine optimization: visibility in the answers of ChatGPT, Perplexity, Gemini, Claude, Copilot and Google's AI Overviews and AI Mode. This checklist covers 51 checks in seven areas: AI crawler access, the SEO foundations AI visibility depends on, how content is structured into quotable passages, entity and authority signals, structured data, direct testing of what assistants say, and measurement. It follows the same 51 points as the Illucrum GEO Audit and works for any site type.

In short: A GEO audit checks that AI crawlers aren't blocked by robots.txt or a firewall, that key pages are indexed and readable without JavaScript, that pages answer questions early in self-contained passages, that the business is defined consistently across the web, that structured data is valid, what AI assistants actually say about the business and its competitors, and whether AI visibility can be measured. All 51 checks run on free tools.


What does this GEO checklist cover?

This GEO checklist covers 51 checks in 7 sections. A classic SEO audit asks whether a page can rank; a GEO audit asks whether a page can be retrieved, quoted and attributed by an AI system. The two overlap less than most people expect: an Ahrefs study of 15,000 queries found that only about 8% of the URLs cited by ChatGPT, Gemini and Copilot ranked in Google's top 10 for the same query (28.6% for Perplexity).

# Section Checks The question it answers
1 AI crawler and retrieval access 1 to 9 Can AI systems reach the site at all?
2 Classic SEO foundations 10 to 17 Are the search basics that gate AI retrieval in place?
3 Passage and chunk structure 18 to 27 Can each section of a page be quoted on its own?
4 Entity and authority 28 to 35 Do AI systems know who the business is, and trust it?
5 Structured data 36 to 40 Is the markup valid and does it remove ambiguity?
6 Prompt visibility testing 41 to 46 What do AI assistants actually say when buyers ask?
7 Measurement 47 to 51 Can the business see its own AI visibility over time?

Three things make GEO different from SEO. AI companies run separate crawlers for search, live fetching and training, and blocking the wrong one removes a site from answers. Retrieval systems compete at the level of passages, not whole pages. And assistants build answers from sources they trust, so presence on review platforms, communities and publications counts as much as the site itself.


What do you need before you start?

You need Google Search Console and Google Analytics 4 access, ideally Bing Webmaster Tools and server or CDN logs, and free accounts on the main AI assistants.

Tool Used for Cost
ChatGPT, Perplexity, Gemini, Claude, Copilot Prompt testing, brand knowledge, cited sources Free tiers
Google Search (private window) and Google AI Mode AI Overview and AI Mode citations Free
Google Search Console Indexing, rankings, the generative AI report and the AI features setting Free, needs a verified site
Bing Webmaster Tools Bing indexing and the AI Performance report (Copilot citations) Free, needs a verified site
Google Analytics 4 AI referral traffic and lead attribution Free, needs property access
Server logs or CDN bot analytics Which AI crawlers actually visit Depends on hosting
Illucrum Bulk SEO Checker What a non-rendering crawler sees: raw-HTML headings, text, links and schema for up to 25 URLs Free, no account
curl, Rich Results Test, Schema Markup Validator Crawler access tests and structured data Free

Paid AI visibility trackers exist, but none is needed for this checklist. Four decisions come before the first check:

  1. Pick 5 to 8 priority pages chosen for commercial intent rather than traffic: service or product pages, pricing, comparisons, key guides.
  2. Write a prompt set of 10 to 15 prompts your buyers would type into an assistant, in their language. Cover discovery ("best [category] for [buyer]"), problem ("how do I fix [problem]"), comparison ("[brand] vs [competitor]"), qualification ("how much does [service] cost") and brand ("what is [brand]"). Keep the set fixed; it is your baseline for rerunning later.
  3. Test logged out or in a temporary chat with memory off, run each prompt two or three times, and record the date. Answers vary between runs.
  4. Open a spreadsheet with four columns: check number, finding, evidence and severity.

Section 1: How do you check AI crawler and retrieval access? (checks 1 to 9)

AI crawler access checks establish whether AI systems can reach the site. Everything else in this checklist depends on them, so work through this section first.

Crawler class What it does Examples Effect of blocking
Retrieval Builds the index an assistant searches when it answers OAI-SearchBot, Claude-SearchBot, PerplexityBot, Bingbot, Amazonbot, DuckAssistBot The site drops out of the candidates for citation
User-triggered Fetches a page because a person asked the assistant to read it ChatGPT-User, Claude-User, Perplexity-User, MistralAI-User The assistant tells a live prospect it can't read the site
Training Collects content that may train future models GPTBot, ClaudeBot, Google-Extended, Applebot-Extended, CCBot, meta-externalagent Little effect on citations today; a consent choice

1. Retrieval crawler access

Open /robots.txt, confirm it returns 200, and read the rules that apply to each retrieval crawler in the table above. Then read the wildcard rule (User-agent: *) on its own: a broad disallow there applies to every crawler without a rule of its own, which is how retrieval crawlers get blocked without anyone intending it. OpenAI says sites that block OAI-SearchBot won't be shown in ChatGPT search answers.

Good result: every retrieval crawler allowed, by name or through a permissive wildcard.

Record: the rule that applies to each retrieval crawler, and whether each block looks deliberate or inherited.

2. User-triggered fetcher access

User-triggered fetchers act because a real person, often a prospect, pasted your URL into an assistant or asked it to read your page. Check the rules for them and confirm they aren't caught by the wildcard. OpenAI and Perplexity both say robots.txt may not apply to these fetchers because a user initiated the request, so a block usually lives in a firewall rule rather than robots.txt.

Good result: user-triggered fetchers allowed in robots.txt and at the firewall.

Record: the rule for each fetcher and any firewall rule that affects them.

3. Training crawler policy

Check the rules for training crawlers and decide whether the current state is a decision or an accident. Blocking training crawlers has little effect on whether AI products cite the site today; it is a consent and licensing choice, and either position is defensible. Google-Extended controls use of content for Gemini training and grounding, and Google says it doesn't affect inclusion or ranking in Google Search.

Good result: a deliberate policy on training crawlers that doesn't also catch retrieval crawlers.

Record: the rule for each training crawler and whether it was a decision.

4. Firewall, CDN and bot protection

Robots.txt states intent; the firewall, CDN or bot-protection layer enforces reality, and the two disagree more often than site owners expect. Find out what sits in front of the site (Cloudflare or another CDN, a hosting bot filter, a security plugin) and check its AI crawler settings. Request a key URL with a retrieval crawler's user agent, for example curl -A "OAI-SearchBot" https://yourdomain.com/, and compare it with a normal request. A 403, a challenge page or a CAPTCHA means the edge is blocking it.

Good result: the edge lets through every crawler robots.txt allows.

Record: the protection in place, its AI crawler setting and the result of each user-agent test.

5. Crawler activity in server logs

If you have no access to server logs or CDN bot analytics, mark this check N/A. Otherwise, pull at least 14 days of logs and filter for AI user agents, for example with the pattern GPTBot|OAI-SearchBot|ChatGPT-User|ClaudeBot|Claude-User|Claude-SearchBot|PerplexityBot|Perplexity-User|Google-Extended|CCBot|Amazonbot|meta-externalagent|DuckAssistBot. Record which agents appear, how often, which URLs they fetch and which status codes they get. Some providers verify crawlers by IP address, so logs are the only proof that the real crawler gets through.

Good result: regular visits from retrieval crawlers, served with status 200.

Record: each agent seen, its visit count, the URLs it fetched and the status codes it received.

6. Google AI features setting and snippet directives

In Search Console, open Settings and check the Search generative AI control, which removes a site's content from AI Overviews, AI Mode and Discover's AI features. Google made it available worldwide in August 2026 and says it isn't used as a ranking signal elsewhere in Search. Then check key pages for nosnippet, a low max-snippet value, and data-nosnippet around meaningful content; these restrict AI features and regular snippets alike.

Good result: no opt-out, and snippet directives that allow full extraction of key pages.

Record: the state of the AI control and any snippet directive on priority pages.

7. Bing indexing and Copilot grounding

Microsoft Copilot grounds its answers in the Bing index, which makes Bing indexing matter more than Bing's own search share suggests. Compare the pages Bing has indexed (site:yourdomain.com on Bing, or Bing Webmaster Tools) with Google's count. Check Bing Webmaster Tools is set up with the sitemap submitted, and check robots.txt for Bingbot rules.

Good result: Bing indexing comparable to Google's, with Webmaster Tools set up.

Record: Bing and Google indexed counts, Webmaster Tools status and any Bingbot rule.

8. Raw HTML parity and JavaScript dependency

Google renders JavaScript; most AI retrieval and user-triggered fetchers largely don't. Paste each priority page into the Bulk SEO Checker, which reads the raw HTML as those fetchers do. If a page that looks complete in the browser comes back with no H1 or a word count near zero, its content depends on JavaScript. Check specifically for pricing tables, tabs, testimonials, FAQ answers and specifications that appear only after scripts run, and for key facts that exist only inside images.

Good result: headings, body copy and commercial facts all in the raw HTML.

Record: per priority page, what is missing from the raw HTML.

9. llms.txt and machine-readable surfaces

llms.txt is a proposed file that lists a site's key content for AI systems. Check whether /llms.txt exists and, if it does, whether logs show any AI agent fetching it. Weight this check low: no major AI search product has confirmed using the file, and Google's John Mueller called it "purely speculative" in June 2026. Its practical use today is developer documentation read by coding assistants. Check that any implementation hasn't created indexable Markdown copies of your pages, which is duplicate content.

Good result: present and harmless, or absent. Absence isn't a fault.

Record: whether the file exists, whether anything fetches it and whether it created duplicate pages.


Section 2: How do you check the classic SEO foundations? (checks 10 to 17)

Classic SEO foundation checks cover the search basics that decide whether a page can be retrieved at all. Good SEO doesn't guarantee AI citations, but weak SEO reliably prevents them: an unindexed page isn't in any retrieval index, a page split across four URLs splits its signals, and a page that times out is skipped.

10. Indexation health

In Search Console, open Indexing, then Pages, and record the indexed and not indexed counts and the reasons. Run URL Inspection on each priority page. A page Google hasn't indexed can't appear in AI Overviews or AI Mode, and is a weaker candidate everywhere else.

Good result: every priority page indexed.

Record: indexed and not indexed counts, reasons, and each priority page's status.

11. Canonicalization and duplicate content

Check each priority page carries a self-referencing canonical tag (the Bulk SEO Checker lists them), and look for the same content at several URLs: parameters, trailing slashes, http and https, www, print versions. Where the same passage exists at several addresses, a retrieval system has no single source to cite.

Good result: each piece of content at one consistent URL.

Record: duplicate URL groups and missing or wrong canonicals.

12. XML sitemap and discovery

Confirm the sitemap is valid, lists only live canonical URLs, is declared in robots.txt (which is how several non-Google crawlers find it) and is submitted to both Google Search Console and Bing Webmaster Tools. Check the lastmod dates are real rather than all set to today.

Good result: a clean sitemap with accurate dates, submitted to Google and Bing.

Record: submission status in each tool, and the state of the lastmod dates.

13. Speed and fetch reliability

Run the priority pages through PageSpeed Insights and record the mobile Performance score and Core Web Vitals. Check server response time (time to first byte) in the DevTools Network tab, and whether rate limiting throttles repeated requests. User-triggered fetchers work while a person waits; a slow response means the assistant reports it couldn't read the page.

Good result: Core Web Vitals in Google's Good range and a fast, reliable server response.

Record: scores, Core Web Vitals and server response time per priority page.

14. HTTPS, status codes and redirects

Confirm HTTPS is enforced with a valid certificate and no mixed content, that priority URLs return 200, and that redirects resolve in one step with permanent (301 or 308) codes. AI crawlers visit less often than Googlebot and are slower to follow content that has moved, and a cited URL that later breaks becomes a broken citation.

Good result: HTTPS everywhere, clean status codes and single-step redirects.

Record: errors on linked or indexed URLs and any redirect chain.

15. Title, description and H1 consistency

Check each priority page has a unique title of about 60 characters or less, a unique meta description and an H1 that states the subject literally. Then check the title, the H1 and the headline in any Article markup describe the same thing. The Bulk SEO Checker lists all three for up to 25 pages at once.

Good result: unique, literal titles and H1s that agree with each other and with the markup.

Record: duplicates, vague H1s and mismatches between title, H1 and headline.

16. Internal linking and topic clusters

A topic cluster is a group of pages on one subject that link to each other and to a main page. Check priority pages are within three clicks of the homepage and not orphaned, that related content is grouped into clusters rather than a flat list of posts, and that anchor text describes the target. Clear clusters help assistants work out what the site is an authority on.

Good result: clear clusters, no orphan priority pages and descriptive anchors.

Record: each priority page's click depth and incoming links, and orphan pages.

17. Organic ranking baseline

In Search Console, record the average position, impressions and clicks for your priority non-branded queries, and which pages rank. Use this as context, not as a measure of AI visibility: ranking and citation overlap only partly, so a strong baseline doesn't prove AI visibility and a weak one doesn't rule it out.

Good result: established rankings for the priority non-branded queries.

Record: position, impressions and clicks per priority query, with the ranking page.


Section 3: How do you check passage and chunk structure? (checks 18 to 27)

Passage structure checks look at how quotable each priority page is. Retrieval systems don't read a page whole: they split it into passages, score each passage against the question and cite the best ones. A page's opening, each section, each FAQ answer and each table row compete separately. The working thresholds below come from how chunking behaves, not from any published rule: answer blocks of roughly 40 to 75 words, paragraphs under roughly 100 words, one idea per paragraph.

18. Answer-first opening

For each priority page, write down the question its H1 implies. Read the first 100 words: do they answer that question directly and completely? Flag pages that open with scene-setting, brand story or a promise of what the page will cover. Check each H2 section the same way: does it open with its answer?

Good result: every priority page answers its question in the first 100 words.

Record: per page, where the direct answer first appears.

19. Self-contained passages

Read each H2 section on its own, as if it arrived with no context. Does it still make sense, or does it depend on a subject or definition from earlier? Flag sections that start with "This means...", "As mentioned above..." or "The second reason is...". Each section should be quotable and attributable without becoming misleading.

Good result: sections that make sense read alone.

Record: sections that depend on earlier text, with the dependent phrase.

20. Paragraph length and one idea per paragraph

Flag paragraphs longer than about 100 words and paragraphs that carry more than one idea or fact. Check key commercial facts (price, process, eligibility, turnaround time) sit in their own short paragraph rather than in the middle of a long one. A paragraph with three separate claims matches no single question well.

Good result: short, single-idea paragraphs, with key facts standing alone.

Record: long or multi-idea paragraphs per page.

21. Literal, descriptive headings

Read the H2 and H3 headings of each priority page on their own, using the Headings table of the Bulk SEO Checker. Can you tell what the page covers from the headings alone? Flag clever, punning, teasing or metaphorical headings. Headings should name their subject and, where it fits, ask the question a user would ask.

Good result: an outline that explains the page on its own.

Record: headings that don't state their subject.

22. Explicit entity naming

A passage retrieved on its own has no surrounding context: "It costs $500" can't be attributed, while "The [Brand] [service] costs $500" can. Look for pronouns and stand-ins ("it", "they", "this", "the company", "our service") carrying the main subject of a section, and check the brand, product and service names appear within each section, not only in the introduction.

Good result: the main entity named in every section.

Record: sections whose subject is only a pronoun or stand-in.

23. Comparison tables and structured lists

Check that pages comparing options (services, plans, alternatives) present the comparison as a table rather than prose, that processes are numbered lists, and that specifications (pricing, inclusions, timelines, eligibility) are tabulated. Confirm tables are real HTML <table> elements with headers, not images or styled blocks.

Good result: comparisons in real HTML tables and processes in numbered lists.

Record: comparisons still in prose, and tables that are images or styled blocks.

24. FAQ blocks in prompt language

Check priority pages have real FAQ blocks, not marketing copy formatted as questions. The questions should be phrased the way a buyer types them into an assistant ("How much does [service] cost?"), not the way marketing writes them ("Why choose us?"). The answers should be complete on the page, not teasers pointing to a contact form. Ask an assistant the same questions to see how it phrases them.

Good result: real questions in buyers' words, answered completely on the page.

Record: per page, whether an FAQ exists, how its questions are phrased and whether answers are complete.

25. Fact density and verifiable specifics

Compare the specific, checkable facts on each priority page (numbers, prices, dates, durations, named tools and standards, examples) with general assertions ("industry-leading", "tailored to your needs", "proven results"). The test: could a competitor put the same sentence on their site unchanged? If so, it identifies nothing and can't be cited. Check claims that could be sourced are sourced.

Good result: pages dense with specific, checkable facts.

Record: per page, the count of specific facts and the generic claims worth replacing.

26. Semantic HTML and hidden content

Check pages use semantic elements (article, section, nav, h1 to h6, table, ul, ol) rather than only divs, with one H1 and no skipped heading levels. Check content in accordions, tabs and modals is in the initial HTML and not loaded on click, that data-nosnippet isn't wrapped around meaningful content, and that no key text exists only in an image.

Good result: semantic markup, a correct heading order and no key content hidden from fetchers.

Record: hierarchy errors, content loaded on interaction and text locked in images.

27. Freshness and date accuracy

Check each page shows a published or updated date, and that datePublished and dateModified in the markup are present and accurate. Flag two opposite failures: dates not updated after real changes, and dates refreshed with no change. Look for outdated facts on priority pages: old statistics, retired tools, past years written as current.

Good result: accurate visible and markup dates, and no outdated facts.

Record: pages without dates, inaccurate dates and outdated facts.


Section 4: How do you check entity and authority signals? (checks 28 to 35)

Entity and authority checks establish whether AI systems know who the business is, trust it and describe it accurately. Assistants cross-check facts across independent sources, so what a site says about itself establishes little on its own. This is also the section least fixable through site work: where the gap is recognition, the honest answer is months of presence-building on other sites.

28. Brand entity definition

Check the homepage for Organization markup, or a specific type, with name, url, logo, description and a sameAs list linking to the brand's real profiles (LinkedIn, Crunchbase, review platforms, directories). Check the About page states plainly what the business does, for whom and where, and that this matches the descriptions on external profiles.

Good result: complete organization markup with sameAs, and a plain definition of the business on the site.

Record: markup fields present, the sameAs list and any description that differs from external profiles.

29. Author entity resolution

Check authors have Person markup with name, url, jobTitle and sameAs links to their LinkedIn and other profiles, and that Article markup references that Person rather than a plain text name. Check the author's name is written identically in the byline, the markup, the author page and on LinkedIn.

Good result: Person markup with sameAs, and one consistent spelling of each author's name.

Record: author markup present, sameAs links and any name variants.

30. Named bylines and credentials

Check every article and guide has a named human byline, not "Admin" or "[Brand] Team", and that the byline leads to an author page with specific, verifiable experience: roles, years, employers, qualifications, published work. A bio earns trust when a reader could check it.

Good result: named authors with specific, checkable bios.

Record: the share of content with a named author and the state of the bios.

31. Consistency across the web

List the brand's external presence: LinkedIn, Crunchbase, the Business Profile, review platforms, directories, partner pages. Compare name, description, services, address and contact details across them. Flag legal names in one place and trading names in another, services no longer offered, old addresses, and abandoned or duplicate profiles. When sources contradict each other, machines trust all of them less.

Good result: the same description and details on every profile.

Record: each profile with inconsistent or outdated details.

32. Knowledge panel and Wikidata

Search the brand name on Google and Bing and check whether a knowledge panel appears and is accurate. Search Wikidata for the brand and its founders. Judge realistically whether the business meets notability thresholds; most small businesses don't, and for them this check is N/A rather than something to pursue directly.

Good result: an accurate knowledge panel or Wikidata entry, or N/A where notability isn't realistic yet.

Record: whether each exists and any wrong information in it.

33. Third-party citation footprint

Start from the list of sources assistants actually cite for your category, which you'll build in check 46. That list defines where presence matters; don't audit directories no assistant cites. For each source, check whether the business appears and how its presence compares with the competitors being recommended, in volume and recency.

Good result: a current presence on the sources assistants cite for your category.

Record: per cited source, your presence and the recommended competitors' presence.

34. Sentiment and accuracy of off-site mentions

Read what the sources from check 33 actually say about the business. Record the sentiment (positive, neutral, negative or mixed), check facts such as services, pricing and category, and flag outdated descriptions and confusion with similarly named businesses. Assistants repeat what these sources say, and most publishers will correct a factual detail on request.

Good result: accurate mentions that are positive or neutral. Mark N/A if there are too few to judge.

Record: each source, its sentiment and any factual error.

35. Original data and citable assets

Check whether the site publishes anything that exists nowhere else: its own data, research, benchmarks, a documented method, tools or worked examples. Restating an industry statistic gives an assistant no reason to cite you over the original. Note data the business holds but hasn't published, such as aggregated outcomes, price ranges or process timings.

Good result: some original material that others would cite.

Record: original assets published, and unpublished data worth publishing.


Section 5: How do you check structured data for AI search? (checks 36 to 40)

Structured data checks look at the markup that removes ambiguity about what a page is and who it comes from. Structured data supports citation but doesn't create it, which is worth remembering because schema is often oversold. Google stopped showing FAQ rich results on May 7, 2026 and removed FAQ support from the Rich Results Test in June 2026, so validate FAQ markup with the Schema Markup Validator instead.

36. Article and author markup

Validate Article or BlogPosting markup on content pages. Check headline matches the H1 exactly, author references a Person and publisher references the Organization, and datePublished and dateModified are present and accurate.

Good result: valid Article markup with a matching headline, a linked author and accurate dates.

Record: errors, headline mismatches and stale dates.

37. FAQPage markup

If no priority page has real question-and-answer content, mark this check N/A. Otherwise, check pages with real FAQs carry FAQPage markup that validates and matches the visible questions and answers exactly, and flag FAQ markup on content that isn't questions and answers. The Google rich result is gone and the AI benefit is unproven, so treat this as low-cost housekeeping.

Good result: valid FAQ markup that matches the page, where real FAQs exist.

Record: pages with FAQs, whether they carry markup and any mismatch.

38. Organization, service and product markup

Check the most specific type is used where one applies (LocalBusiness, ProfessionalService, LegalService, SoftwareApplication), that service pages carry Service markup and product pages Product with offers, and that prices in the markup match the page. Check areaServed, serviceType and contact details where they apply.

Good result: specific types with complete fields and prices matching the page.

Record: types used per page type and missing or mismatched fields.

39. Breadcrumb markup

Check interior pages carry BreadcrumbList markup, that the trail matches the site's real hierarchy, and that the same trail is visible on the page. Breadcrumbs state where a page sits in the site's topics, which supports the clusters from check 16.

Good result: valid breadcrumb markup matching the real structure, on every interior page.

Record: templates without breadcrumbs and trails that don't match the structure.

40. Schema validity and content match

Run every priority page through both validators, or paste them into the Bulk SEO Checker to list every block and catch any that fail to parse. Check for structural failures (malformed JSON-LD, missing @context), claims in the markup that the page doesn't show (prices, ratings, dates, authors), and duplicate or conflicting blocks, often from a theme and a plugin both adding markup.

Good result: all markup valid and matching the visible content, from a single source.

Record: errors, mismatches and conflicting blocks per page.


Section 6: How do you test AI prompt visibility? (checks 41 to 46)

Prompt visibility tests measure whether citation is actually happening. Everything before this section checks the conditions for it. Use the fixed prompt set from your setup, run each prompt two or three times per platform, and record the date and the pattern rather than a single answer.

41. Brand knowledge and accuracy

Ask each assistant "What is [brand] and what does it do?", then "What services does [brand] offer and what do they cost?" and "Who runs [brand]?". Record whether the brand is recognized, whether the description is accurate, any invented details and the sources cited. Confident inaccuracy is worse than absence, because a prospect gets wrong information with no warning.

Good result: recognized and described accurately by every assistant tested.

Record: per assistant, the description, any errors and the cited sources.

42. Unprompted recommendations

Run the discovery and comparison prompts without naming the brand. Record whether it is recommended, in what position and how it is described, and which competitors appear consistently. This is the headline measure. A brand that is recognized when named (check 41) but not recommended when unnamed usually has thin third-party presence rather than a content problem.

Good result: recommended for category prompts by two or more assistants.

Record: per assistant and prompt, whether you appear, your position and the competitors named.

43. Google AI Overview citations

Run the prompts as Google searches in a private window. Record which trigger an AI Overview, whether your site is cited and which domains are. AI Overview citation and ranking are only partly linked: a page below the fold can be cited and the first result can be left out. Seer Interactive found brands cited in an AI Overview had 35% higher organic CTR than those that weren't.

Good result: cited in AI Overviews for several priority queries (or no Overviews appear).

Record: per query, whether an Overview appears, the cited domains and whether you are among them.

44. Google AI Mode presence

AI Mode is Google's conversational search mode, which handles multi-step questions differently from AI Overviews. If it isn't available in your market, mark this check N/A. Run the prompts in AI Mode and record citations, including on follow-up questions, which are usually more specific. Compare with check 43; differences between the two are common.

Good result: cited across priority questions and their follow-ups.

Record: per prompt, citations on the first answer and on follow-ups.

45. Competitor citation share

Across every prompt tested, count how often each competitor is cited or recommended, and calculate your share of all citations. Check whether the most-cited competitors are the same businesses you compete with in regular search; often they aren't, which is itself a finding. For the top two or three, note what sets them apart: review volume, community presence, press coverage, original data or structure.

Good result: a citation share comparable to the named competitors.

Record: citations per brand across the prompt set, and what distinguishes the leaders.

46. Cited source analysis

List every domain the assistants cited across all prompts, not only those mentioning you, and rank them by frequency. Group them: review platforms, communities, publications, competitor sites, directories, independent blogs. Cross-check with check 33 for where you appear. This turns "build authority" into a ranked list of the specific places where presence changes answers.

Good result: your brand present on the most frequently cited sources.

Record: the ranked source list, grouped, with your presence marked on each.


Section 7: How do you measure AI search visibility? (checks 47 to 51)

Measurement checks establish whether the business can see its own AI visibility after the audit. Every measurement source is partial: Google's report shows impressions only, analytics misses AI visits that arrive without a referrer, and nothing captures citations that never produce a click. The fixed prompt set from your setup remains the most complete signal.

47. Search Console generative AI report

Check whether the property has Search Console's generative AI performance report, which Google introduced in June 2026 and made available worldwide in August 2026. Record impressions in AI Overviews and AI Mode by page, country and date; the page breakdown shows which content Google's AI features draw on. Know its limits: impressions only, with no clicks, CTR or query data.

Good result: the report available, reviewed, with AI impressions recorded by page.

Record: AI impressions for the top pages and the date range.

48. Bing AI Performance report

Check Bing Webmaster Tools is set up and whether its AI Performance report, launched in public preview in February 2026, shows data. Record citations, cited pages and the grounding queries that triggered them for Copilot and Bing's AI answers. Unlike Google's report, it shows which questions your content answers.

Good result: Webmaster Tools set up, with the AI Performance report reviewed.

Record: citation count, top cited pages and top grounding queries.

49. GA4 AI traffic segmentation

If you have no Analytics access, mark this check N/A. In GA4, check whether the AI Assistant channel, added to the default channel group in May 2026, shows sessions. Google named ChatGPT, Gemini and Claude but hasn't published its full source list, so check where Perplexity and Copilot sessions land and cover gaps with a custom channel group ordered above Referral. AI visits without a referrer, often from mobile apps, still land in Direct.

Good result: AI traffic fully segmented, with the gaps covered.

Record: sessions per AI source, where uncovered sources land, and the landing pages AI visits arrive on.

50. AI crawler monitoring

If you have no access to crawler activity data, mark this check N/A. Otherwise, check logs or CDN bot analytics are kept long enough to show trends and that someone looks at them. A drop in retrieval crawler visits shows up before any drop in citations; CDN configuration changes and security plugin updates are common causes.

Good result: crawler activity retained and reviewed regularly.

Record: retention period, review routine and any unexplained drop.

51. Lead source attribution

Where AI traffic is segmented (check 49), check whether a lead from an AI assistant stays identifiable when the visitor submits a form, for example through a hidden first-touch source field carried into your CRM. Compare AI-referred conversion with organic search conversion against your own baseline rather than a published benchmark, because published figures vary widely by site and method.

Good result: the AI source captured from first visit through to the CRM.

Record: where the source is kept and where it is lost.


Frequently asked questions

What is the difference between GEO and SEO?

SEO makes pages rank in search results; GEO makes pages retrievable and quotable in AI answers. GEO depends on SEO foundations such as indexing and speed, but adds checks SEO doesn't have: AI crawler classes, passage structure, off-site entity signals and direct prompt testing. The GEO audit vs SEO audit comparison covers when each one fits.

Does llms.txt help with AI search visibility?

There is no evidence that llms.txt helps with AI search visibility today. No major AI search product has confirmed using the file for retrieval or ranking, and Google's John Mueller called it "purely speculative" in June 2026. It has a practical use for developer documentation read by coding assistants, and it does no harm if it doesn't create duplicate pages.

How long does a GEO audit take to do yourself?

A first pass through this GEO checklist takes six to ten hours. The access checks in section 1 take under an hour. The prompt testing in section 6 takes the longest, because each prompt runs two or three times on several platforms, and the passage structure review takes an hour or two per handful of priority pages.

Should a small business work on GEO before SEO?

Usually not. AI citations mostly come from sites that are already indexed, reasonably authoritative and mentioned on trusted third-party sources. For a site with weak SEO basics, the foundations in section 2 and a checklist such as the website SEO checklist come first. The exceptions are the quick wins in section 1: an AI crawler block is worth fixing today.

How often should you check AI visibility?

Rerun the fixed prompt set from this GEO checklist monthly or quarterly, on the same platforms, and compare with your first results. Check Search Console's generative AI report and Bing's AI Performance report monthly. A full GEO audit once a year, or after a site migration or CDN change, covers the rest.


AI visibility starts with access and ends with reputation: crawlers that can get in, pages that answer early in passages that stand alone, and a business that other trusted sources describe accurately. If you'd rather have the whole picture gathered and prioritized for you, the GEO Audit covers the same 51 points at a fixed price listed on the pricing page.