The First Ninety Days With An AI SEO Agency
Two wrong answers circulate about how long this takes. One says a few weeks, which sells engagements and then disappoints. The other says a year or more, which is used to defer starting and to excuse a lack of movement halfway through.
On Third Party Tracking Tools Several tools now offer to monitor this at scale, and they save real time once your prompt set runs into the hundreds. They are worth buying for trend lines and for coverage you cannot manually sustain.
Include one deliberately open question at the end, asking what they would do differently from what the brief proposes. A good supplier will disagree with something, and the disagreement is worth more than the rest of the proposal because it shows they read the situation rather than the request. A response that agrees with every assumption in your brief has told you nothing you did not already believe.
Month Two: Corrections and the First Rewrites The work should now be concentrated on the recurring sources from the baseline. Expect a list of listings claimed, details corrected and errors submitted, with names and dates attached.
The truthful answer is that different parts of the work move at very different speeds, and knowing which is which lets you judge an engagement at the right moment instead of the convenient one. llm visibility tracking
Also decide up front who owns this. Measurement that belongs to everyone gets run inconsistently, the conditions drift, and the series becomes uncomparable within two quarters. One named person running a modest set reliably produces more usable information than a sophisticated programme with no owner.
That sequence typically takes a few weeks per source, and the effect on answers follows once enough of the recurring sources agree with each other. This is the phase where identity work begins to pay, and it is slower than people expect because it depends on other people's publishing schedules.
Results Split by Intent, With Run Counts Not one number. Mention rate reported as a fraction with the run count visible, broken out by prompt tier, so buying intent is never blended with definitional questions.
Tracking this is genuinely awkward, and pretending otherwise is how most reporting in this field goes wrong. There is no console. Answers vary between runs. Referral attribution is inconsistent between assistants. Anyone handing you a single confident number has hidden a great deal of variance behind it.
Fix the Prompt Set and Never Casually Change It Your prompt set is the instrument. If you adjust it between runs you are measuring your own edits, and any trend line you draw afterwards is meaningless.
The monthly report is where an engagement is either accountable or theatrical, and the difference is visible from the first page. A useful report can be argued with. A padded one cannot, because there is nothing in it specific enough to disagree about.
What Padding Looks Like Screenshots of favourable answers with no indication of how many runs produced them. Industry news summaries that could have been written without opening your account. A rising score with no methodology. Traffic charts from unrelated channels included to fill space.
A weak brief produces a generic proposal, and a generic proposal produces a generic engagement that spends the first two months discovering things you already knew. The brief is the cheapest lever you have over the quality of the work.
How to Judge Progress at Each Stage Use different measures at different points rather than asking for mentions from month one. At the end of month one, ask whether the baseline exists and whether access problems were found. At month three, ask whether listings are corrected and whether your own pages appear in citation lists at all.
Control the Session Conditions Personalisation quietly corrupts this. Run from a signed out session, or a fresh session with memory and history disabled, and do not use an account that has been researching your own company all week.
Include the constraints too. The job size you turn down, the sector you do not serve, the situation where a competitor is genuinely the better answer. Those are the statements that get quoted, and an agency will not invent them for you.
State who signs off, how fast, and what is off limits. An agency that knows the constraints will build a plan that fits them. One that finds out gradually will spend the retainer producing work that never ships. llm visibility tracking
One thing that reliably compresses the timeline is starting the slow work first. Outreach and coverage take months regardless of what else is happening, so beginning them in week one rather than month four moves the whole programme forward by a quarter at no additional cost. Most plans do the opposite, sequencing the slow work last because it is the least certain.
Three to Nine Months: Earned Coverage The slowest and most valuable part. Getting into the comparison articles, trade publications and community discussions that assistants actually cite depends on other organisations deciding to write about you, which no amount of budget reliably accelerates.