What An AI SEO Agency Should Report Every Month
It is also worth checking which assistant your customers actually use rather than assuming. The answer varies by profession, age and country far more than industry commentary suggests, and several businesses have built measurement programmes around a system their buyers never open. Adding one question to your enquiry form settles it in a fortnight and can redirect the whole effort.
What a Defensible Business Case Looks Like It states what cannot be measured. It reports inputs completed, with counts. It reports prompt set movement as fractions with visible run counts, split by intent. It includes the soft signals as anecdote clearly labelled as anecdote. It attributes every external statistic.
Refusals matter. A report that only contains successes is either describing a suspiciously easy month or omitting the parts that did not work, and the omitted parts are usually where the useful information is.
It also appears more conservative in commercial categories, hedging or declining to make a direct recommendation more often than the others. Where it does recommend, established entity signals seem to matter, which favours brands with consistent details and long records over newer entrants.
Crawler access restored on a date. Listings claimed and corrected, with a count. Factual errors fixed on third party sources, with a count. Pages published that answer prompts your baseline showed were being answered badly. Reviews responded to.
You will find your own category's pattern, which frequently contradicts the general one. Some industries are dominated by a single trade directory. Others are dominated by one forum. That specific finding is worth more than any general description of how these systems behave.
Measure Position Change in the Prompt Set This is the closest thing to an output metric that you can genuinely audit, because you own the instrument. Run a fixed prompt set on a fixed schedule under fixed conditions, and track four things:
The exception is a category where assistant use at the research stage is already heavy and where the incumbent comparison pages are weak. There the newer channel can be underpriced, and moving early is worth more than it will be in two years.
Being the Source Instead of the Casualty The summary cites sources, and being one of them is now a legitimate objective. The requirements resemble what earns citations anywhere else: a page that answers directly, contains specifics worth attributing, and is reachable and readable by a crawler.
This variability is the main practical trap. Testing without web access and concluding you are invisible measures the training corpus rather than current retrieval, and the two can disagree sharply. Record which mode you used with every run.
You also cannot cleanly attribute a purchase to a recommendation the buyer received three weeks earlier in a conversation you never saw. That influence is real, it is often the main value of the channel, and it will not appear in any report you own.
Two implications follow regardless of which system you are studying. Being findable by the underlying search step is necessary, and being worth quoting once fetched is what decides whether you are used. Almost everything actionable sits in those two requirements.
There is a defensible way to measure this. It produces less certainty than a paid media report and considerably more than a visibility score, and it has the advantage of surviving scrutiny. ai seo company
How to Split the Budget For most businesses, organic search still delivers the larger share of traffic, so the sensible default is to keep the majority of effort there and carve out a defined share for the newer channel rather than gambling the lot.
Where the Work Is Genuinely the Same The foundations do not change. Crawlable pages, sane site structure, fast rendering, accurate structured data, internal links that reflect how topics relate, and content that answers a real question all serve both channels.
The discipline is in how you report their output. Every one of them samples: their own prompt set, their own infrastructure, their own run frequency. Their number is an estimate from a particular vantage point, not a count of what happened.
Work Completed, in Countable Units Listings claimed, with names. Errors corrected, with the source and what was wrong. Pages published or rewritten, with URLs. Technical changes made, with dates. Outreach attempted and its outcome, including refusals.
What Padding Looks Like Screenshots of favourable answers with no indication of how many runs produced them. Industry news summaries that could have been written without opening your account. A rising score with no methodology. Traffic charts from unrelated channels included to fill space.
Gemini and Google Surfaces Closest to conventional search infrastructure, which has a practical consequence: work that improves your standing in Google search tends to carry over here more than it does elsewhere.