AI SEO Services That Move Revenue, Not Vanity Metrics
Good versions read like this: mention rate on evaluation prompts rose from two in fifteen to six in fifteen, which we attribute to the three directory corrections completed in week two, though a competitor also stopped publishing during the same period.
The problem is not that the tools are dishonest. It is that the vendor controls both the number and the prompt set that produces it, so the score can improve without anything happening to your business, and a client has no way to audit the difference.
One test separates a report written to inform from one written to reassure. Read it and try to write down a question it does not answer. In a good report you will find several, because it contains enough specifics to make new questions obvious. In a padded one you will struggle, not because everything is covered but because there is nothing specific enough to interrogate.
The last of these is the most common and the hardest to see, because it produces no error anyone internally encounters. Your site works perfectly in every browser while returning a challenge page to every legitimate retrieval agent.
It is also worth checking whether you are being confused with somebody else rather than ignored. Short names, generic names and names that begin with a number collide with other organisations more often than distinctive ones. Where that is happening, the answer will contain facts that are true about a different company, which reads as a hallucination and is usually an identity collision with a specific fixable cause.
The Prompt Set, Unchanged The report opens with the prompt set used, versioned and dated, and a statement that it is identical to last month's. If it changed, the change is listed explicitly with a reason, and the previous series is kept alongside so comparisons remain honest.
The monthly report is where an engagement is either accountable or theatrical, and the difference is visible from the first page. A useful report can be argued with. A padded one cannot, because there is nothing in it specific enough to disagree about.
Fair Reasons for Flat Results Not every flat quarter is a failure, and being unfair about this loses good suppliers. A saturated category takes longer. A site that needed substantial technical work will have spent the first months on it. Earned coverage depends on other organisations publishing, which nobody can schedule.
Watch the quality of enquiries as well as the count. A common early signal is that conversations start further along, with the prospect already aware of your price band, your typical timeline and what you do not do, because a machine told them before they arrived. That shows up in sales cycle length and in fewer wasted calls long before it shows up in any dashboard.
What Honest Reporting Contains The prompt set, versioned and unchanged since last month. The raw answers, kept in full rather than summarised. Which competitors were named. Which sources were cited. What work was done. What moved, and the specific claim about which work caused it.
The decision that almost never makes sense for a commercial business is blocking the agents that fetch pages when composing answers. That is the mechanism by which you get recommended, and turning it off is the equivalent of declining to be listed anywhere, taken quietly, usually by accident.
This entire area usually amounts to a day of work. It is routinely the difference between a brand that appears in answers and one that does not, and it is worth doing before anybody writes a single word of new content. answer engine optimization
What llms.txt Proposes It is a proposed convention: a file at your root offering a curated, plain text guide to your site for language model consumers, pointing at the documents you consider authoritative.
Check your robots file, then check your server logs for the relevant agents and see what status codes they receive. A site that returns a challenge to every non-browser request is invisible to this entire channel, and nobody involved will have thought of it as a marketing decision.
One check is worth running independently once a quarter, without telling anyone. Take ten prompts from the agreed set, run them yourself in a signed out session, and compare what you find against the most recent report. Broad agreement is reassuring. A consistent gap in the agency's favour is the single most informative finding available to you, and it is not something a report will ever surface.
Ask for a Small, Bounded Commitment Do not ask for a year. Ask for one quarter with a defined scope: run the baseline, fix access problems, correct the listings on the sources that appeared, publish two pages that answer the questions your baseline showed were answered badly.
It is also worth recording the reason for every rule you keep. A disallow line with no explanation gets preserved indefinitely through migrations and redesigns because nobody dares remove something they do not understand. A one line comment saying who added it and why turns a permanent mystery into a decision that can be revisited.