An Honest Look At AI SEO Agency Pricing Models
What Not to Do in the Name of Legibility Hidden text intended only for machines fails on every axis. It is detectable, it violates most guidelines, and it produces exactly the uniform low quality signal you were trying to avoid.
Report frequency rather than presence. Being named in one run out of five is a genuinely different situation from being named in five out of five, and a report that collapses both to mentioned has thrown away the useful part.
Fix the Prompt Set and Never Casually Change It Your prompt set is the instrument. If you adjust it between runs you are measuring your own edits, and any trend line you draw afterwards is meaningless.
And do not let anyone rewrite your entire site in the flat, listicle heavy register that is currently fashionable in this discipline. It reads as machine assembled to human beings, and content that reads that way tends to be treated as low quality by both audiences.
The honest position is that attribution in this channel is harder than in any other you are currently running, and the field has responded to that difficulty mostly by inventing numbers. Confident figures circulate widely, and a surprising share of them trace back to a vendor's own sample or to a study far smaller than the claim implies.
So attribute it by name every time it appears in a report. A visibility figure presented without saying which tool produced it and how it was sampled will eventually be quoted back at you as fact by somebody who did not know it was an estimate, and that is a difficult correction to make in front of a board. geo seo agency
Build the run into an existing routine rather than creating a new one. Measurement programmes in this field fail through quiet abandonment rather than through a decision, and a modest set attached to an established monthly process survives far longer than an ambitious one that depends on somebody remembering to start it.
Test it rather than assuming. Load your key pages with JavaScript disabled and see what survives. If the product specifications, pricing, service areas and contact details vanish, that is what a machine reads.
Control the Session Conditions Personalisation quietly corrupts this. Run from a signed out session, or a fresh session with memory and history disabled, and do not use an account that has been researching your own company all week.
A false trade off gets invented early in most of these projects. Somebody proposes stripping the design, flattening the copy and restructuring everything around what a crawler finds convenient, and somebody else correctly points out that this would make the site worse for customers.
The Rendering Question This is the one real technical constraint. Content that only exists after JavaScript executes may be invisible to a retrieval fetch, which is not a browsing session and does not always run scripts.
What a Defensible Business Case Looks Like It states what cannot be measured. It reports inputs completed, with counts. It reports prompt set movement as fractions with visible run counts, split by intent. It includes the soft signals as anecdote clearly labelled as anecdote. It attributes every external statistic.
How to Handle Published Statistics Every figure you repeat should carry its publisher, sample size and date. This is not pedantry, it is self protection, because figures in this field get repeated until nobody remembers the sample.
Measure Position Change in the Prompt Set This is the closest thing to an output metric that you can genuinely audit, because you own the instrument. Run a fixed prompt set on a fixed schedule under fixed conditions, and track four things:
Write it once, covering the category question, the problem question, the comparison question, the competitor question and the branded question. Fifty is a workable minimum. Then freeze it, and if you must add prompts later, add them as a separate cohort so the original series stays comparable.
Testing too rarely means you find out about a problem a quarter after it started. Testing too often means drowning in variance that looks like signal and reacting to noise. Both failures are common and the second is more expensive, because it produces work.
Equally, do not publish a stripped alternate version of your site for crawlers. Serving different content to machines than to people is cloaking, it has been penalised for two decades, and there is no reason to expect a more forgiving treatment here.
When to Test More Often Three situations justify a tighter loop. During an active campaign where you need to attribute a specific change, weekly runs on a subset of prompts are reasonable, provided you accept the variance.
What analytics cannot tell you is how often you were named without a click, which in this channel is most of the time. A recommendation that a buyer acts on three weeks later leaves no trace in any report you own. This is why the manual prompt set is not optional, and why nobody should be asked to justify this work on referral traffic alone.