Measurement & Brand

Five Items A Usable Baseline Must Contain

A usable baseline covers five items: crawl-layer state, answer-side mentions and citations, organic position-band distribution, AI-channel visits and enquiries, and branded search volume.

XstraStar editorial team mark

XstraStar Editorial Team

GEO Research & Strategy

Five Items A Usable Baseline Must Contain
In this article
  1. What each item records
  2. Why item 3 cannot be dropped
  3. Every item carries its method
  4. Three traps
  5. How long a baseline takes

A usable baseline covers five items: crawl-layer state, answer-side mentions and citations, organic position-band distribution, AI-channel visits and enquiries, and branded search volume. Drop one and a class of change becomes unexplainable at quarter-end.

A baseline exists so that "how much did it change" has an answer. Without one, quarter-end reporting falls back on absolute values, and absolute values carry no causation.

What each item records

#ItemWhat to recordSource
1Crawl-layer stateRetrieval-crawler request counts and response statuses, whether the server returns complete HTML, whether structured data existsServer logs plus fetch tests
2Answer-side mentions and citationsMention rate on a fixed question pool (daily average), average position when mentioned, whether the own domain appears in cited-domain listsOwn collection or a monitor, plus human verification
3Position-band distributionHow many pages sit in 1–3 / 4–10 / 11–20 / 21+ for target keywordsSearch console or third-party tools
4AI-channel visits and enquiriesSessions and enquiries from AI channels, split by channelSite analytics, server logs, form records
5Branded search volumeBranded search volume and clicks, separated from non-brandedSearch console

Why item 3 cannot be dropped

Organic clicks follow ranking. A third-party per-keyword export (2026-06, United States, one leading site in this category, 4,074 non-branded keywords) shows 13.1 monthly clicks per keyword at positions 1–3, 3.4 at 4–10, 0.1 at 11–20 and 0.0 beyond 21.

Without a starting band distribution, flat sessions at quarter-end cannot distinguish two situations: pages never moved, or pages climbed from beyond 21 to position 15, which is real progress invisible in clicks. Band distribution is the only metric that shows progress at that stage.

Every item carries its method

A baseline is a set of figures with methods attached. Each item states at least four things:

collection window (start and end dates)
sample size (questions × runs per question; or keywords, or pages)
platform and region
algorithm (denominator, de-duplication, treatment of gaps)

⚠️ A baseline without methods fails at the first comparison. The two most common failures: runs per question changed (one run first time, three the second, and the percentages stop being comparable) and the question pool changed (two periods with different questions cannot be compared).

Three traps

Recording an incomplete run as zero. A failed collection day is a gap. Counting it as zero drags the average to a fraction of the real value, so the baseline needs a separate column for valid collection days.

Using a cumulative basis. Cumulative coverage trends toward 100% as the window grows, so a cumulative baseline makes every later comparison show improvement. Use a daily average.

Measuring only the subject. With no comparison set, quarter-end cannot separate "this programme worked" from "the whole category rose". Include how comparable parties performed on the same question set, recorded as distribution and without naming them.

How long a baseline takes

About two weeks, in parallel with the foundations phase. Those two weeks produce three verifiable things: logs proving crawling works, five items archived with methods, and a page matrix yielding an output ceiling. Missing any one of them turns later judgement into a bet.


Key takeaway: A usable baseline covers five items: crawl-layer state (retrieval-crawler requests and statuses in server logs, complete HTML, structured data), answer-side mentions and citations (daily-average mention rate on a fixed pool, average position, own domain present in cited-domain lists or not), organic position-band distribution across 1–3 / 4–10 / 11–20 / 21+, AI-channel visits and enquiries split by channel, and branded search volume kept separate from non-branded. Band distribution is mandatory, because third-party data shows 0.1 monthly clicks per keyword at positions 11–20 and 0.0 beyond 21, leaving band movement as the only visible progress at that stage. Each item states collection window, sample size, platform and region, and algorithm. Three traps: recording failed collection days as zeros, using a cumulative basis, and measuring only the subject with no comparison set. Allow about two weeks, in parallel with foundations.

Related: Baseline before optimisationMeasurement disciplineAI visibility auditFirst-party attributionBranded search lift

Sources: Position bands and click distribution from a third-party keyword tool export (2026-06, United States, one leading site in the category, 4,074 non-branded keywords); estimates, not an acceptance baseline. The five-item list and method requirements from XstraStar delivery practice.

Last updated: 2026-08-27

Turn insight into growth

Find your next AI search growth opportunity

From brand visibility and citation sources to content strategy, the XstraStar team will help you map a clear, measurable GEO optimization path.

Talk to a GEO advisor
XstraStar editorial team mark

About the author

XstraStar Editorial Team

GEO Research & Strategy

The XstraStar editorial team studies AI search, generative engine optimization, and brand visibility, turning platform mechanics, field experience, and market shifts into practical growth guidance.

Related insights

Keep exploring GEO