GEO / Roadmap / AI Citations

The 90-Day GEO Roadmap: How to Test, Measure and Grow AI Citations

A working calendar for a marketing team that wants ChatGPT, Copilot, Perplexity, Gemini and AI Overviews citing its pages within a quarter. It is the companion to our AEO experimentation roadmap, which covers experiment design; this one is the calendar for citation testing.

By AI Studio Team · August 2026 · 11 min read

Quick answer: A 90-day GEO roadmap runs in three phases. Days 1–30 fix the entity, schema and crawlability, pick three intent clusters and publish the first four answer-first pages. Days 31–60 test framing, format and freshness against citation data in Bing Webmaster Tools. Days 61–90 scale the winning framings, broaden the query base and add third-party corroboration. AI Studio measures all of it in citations, not traffic.

On this page

  1. What should you measure before day 1?
  2. Days 1–30: the foundation phase
  3. Days 31–60: the testing phase
  4. Days 61–90: the scaling phase
  5. Week-by-week checklist
  6. Which KPIs matter, and which mislead?
  7. What AI Studio saw in its own data
  8. Frequently asked questions

What should you measure before day 1?

Start with a baseline you can compare against at day 90, or the quarter becomes anecdote. It takes one afternoon and no paid software.

Our guide to tracking AI citations across Bing, Search Console and GA4 covers each source in detail.

Days 1–30: what does the foundation phase include?

The first month has one job: make the site something an assistant can crawl, understand and attribute, then give it four pages worth citing.

Weeks 1–2: entity, schema and crawlability

Weeks 3–4: three clusters, four pages, one scoreboard

Then stop changing things. Engines need a stable site to re-evaluate; a team that keeps editing during the read window cannot tell what worked.

Days 31–60: how do you test what earns citations?

Month two is about hypotheses, not volume. You have four pages and a few weeks of citation data; the question is which properties of those pages the engines respond to.

Four hypotheses worth testing first

What a valid test looks like with noisy data

Citation counts swing week to week for reasons unrelated to you: a model update, a change in report aggregation, a single query spiking. So a valid test changes one variable, runs at least three weeks, and is judged on three numbers: total citations, distinct cited pages, and top-query share. A "win" that only moved the total is suspect. For the discipline of forming and controlling hypotheses, use the AEO experimentation roadmap.

Reading the grounding-query report

The grounding-query view is the closest you get to the assistant's intent. Pages cited for queries matching their intent are winners: reuse the phrasing. Pages cited for queries they were not written for are mis-framed: rewrite the opening and H2s to match. Pages never cited after four weeks indexed are losers: kill or merge them, and write no more in that framing.

Days 61–90: how do you scale what worked?

Month three doubles down on the framings that earned citations and deliberately spreads the risk.

Our GEO strategy guide for Singapore brands covers where Copilot, Gemini and AI Overviews behave differently.

Week-by-week checklist

Weeks are indicative; the sequence is not.

WeekPhaseDoMeasure
0BaselineRecord citations, cited pages, grounding queries; 10 buying questions; entity checkDay-0 snapshot saved
1FoundationCanonical entity description; Organization and Article schema validatedZero schema errors
2FoundationRobots, sitemap, Bing property, llms.txt; fix broken or JS-only pagesKey pages indexed in Bing
3FoundationChoose three intent clusters; draft four answer-first pagesDrafts match the buying questions
4FoundationPublish the four pages; request indexing; start the weekly scoreboardFirst weekly row filled
5TestHold the site stable; read first citations and grounding queriesWhich pages were retrieved, for what
6TestLaunch format and framing tests (one variable each)Test pages indexed
7TestLaunch freshness and hub-and-spoke testsPer-page citations on the scoreboard
8TestRead results at three weeks; kill or merge pages never retrievedWinners and losers named
9ScalePublish the winning framing in the other two clustersCited pages rising
10ScaleAdd adjacent-question pages and a fourth clusterTop-query share falling
11ScaleThird-party corroboration outreach; internal links from hub, siblings, service pageNew referring pages live
12–13ReportRe-run the 10 buying questions; write the day-90 report against the baselineCitations, cited pages, spread, hypotheses won

Takeaway: four weeks of engineering and writing, four of patience and reading, five of deliberate repetition.

Which KPIs matter, and which ones mislead?

Four numbers tell you whether the roadmap is working; one popular number tells you almost nothing in 90 days.

The number that misleads is traffic from AI referrers. Several assistants pass no referrer, so GA4 undercounts it, and it lags citations by weeks or months. Report it as a floor. If leadership wants one chart, give them cited pages over 13 weeks, with the honest caveat that citations are visibility, not leads. Our guide to the best AEO tools in 2026 covers the point at which paid trackers add something the free reports do not.

What AI Studio saw in its own data

We ran a version of this roadmap on our own site in mid-2026. Two caveats: these are Bing Copilot citations as reported in Bing Webmaster Tools, not traffic or leads, and they are one site in one market.

The full write-up is in what 38,000 AI citations taught us. If you would rather have it run for you, our GEO Singapore and AEO agency Singapore pages describe how we scope it and what we report at day 90.

Frequently Asked Questions

Is 90 days long enough to see AI citations move?

Yes, if the foundation is done in the first month and the site is then left stable long enough to be re-crawled. In our own data the first citations appeared within days of publishing a batch of guide pages. What 90 days will not give you is a settled trend, so judge the quarter on cited pages and query spread, not a single peak.

Do we need paid AI-visibility tools to run this roadmap?

No. The roadmap runs on Bing Webmaster Tools' AI Performance report, Search Console's generative-AI views where available, and a spreadsheet of prompts you test by hand. Paid trackers earn their keep when you need daily multi-engine sampling across many prompts, which is usually a day-61 decision rather than a day-one one.

What if our citations are flat after the first 30 days?

Treat it as diagnostic rather than failure. First confirm the new pages are indexed in Bing and Google. Then check the grounding-query report: if your pages are retrieved for queries that do not match their intent, the framing is wrong. If nothing is retrieved at all, the problem is usually crawlability or entity clarity, and days 31–45 should fix that before any new content goes out.

How is this different from the AEO Experimentation Roadmap?

The AEO Experimentation Roadmap is about experiment design: how marketers form hypotheses, control variables and read conversion outcomes. This roadmap is the calendar for a GEO programme whose unit of measurement is the citation. Use that one to design a clean test; use this one to decide what to do in each of the 13 weeks and what to report at day 90.

Can we run this alongside our existing SEO programme, or does it compete?

It complements it. The entity, schema and crawlability work in days 1–30 is the same work a careful SEO team does, and answer-first guide pages tend to earn organic rankings as well as citations. The one real conflict is cadence: GEO testing needs the site held stable during a read window, so schedule redesigns, migrations and mass rewrites around the test calendar.

Should we measure traffic from ChatGPT and Perplexity instead of citations?

Not as the primary KPI. AI referral traffic is undercounted because several assistants pass no referrer, and it lags citations by weeks or months. It is a useful secondary signal, and a lead that names an assistant as its source is worth recording. But the number a roadmap can actually move within 90 days is citations, so build the scoreboard on that.

Related reading

Want to be the answer, not just a search result?

AI Studio builds AI search visibility for Singapore brands — entity, schema, and the content that assistants actually cite. Start with a free AI Visibility Audit of your own site.

Chat on WhatsApp
Book Appointment WhatsApp