How Perplexity, ChatGPT And Gemini Pick Their Sources
Observed behaviour leans toward breadth, pulling from a wider set of sources per answer than the others, and it cites forums, documentation and niche trade sources readily. It also appears comparatively responsive to freshness.
On Third Party Tracking Tools Several tools now offer to monitor this at scale, and they save real time once your prompt set runs into the hundreds. They are worth buying for trend lines and for coverage you cannot manually sustain.
If you must change the prompt set, add new prompts as a separate cohort and keep the original series running unchanged. Editing the instrument retrospectively destroys the comparison you have been building.
The most citable content most businesses could publish already exists, unwritten, in sales calls and support tickets. It is the set of questions people actually ask, with the answers your team gives verbally every week and has never put on a page.
The skill is knowing to sort cited domains by frequency, recognise which of them can be influenced, and understand that a competitor appearing in an answer is usually a story about a third party page rather than about their website. That is a different analytical habit from the one search built.
The specific damage is that somebody sees a dip, rewrites a page, sees the number recover for unrelated reasons, and concludes the rewrite worked. That false lesson then gets applied elsewhere. A slower cadence with more runs per prompt is more informative than a faster one with fewer.
How to Judge Progress at Each Stage Use different measures at different points rather than asking for mentions from month one. At the end of month one, ask whether the baseline exists and whether access problems were found. At month three, ask whether listings are corrected and whether your own pages appear in citation lists at all.
Write it once, covering the category question, the problem question, the comparison question, the competitor question and the branded question. Fifty is a workable minimum. Then freeze it, and if you must add prompts later, add them as a separate cohort so the original series stays comparable.
Record the conditions alongside the results: which assistant, which model version if visible, whether web access was on, the date and the run number. When a result changes sharply, the conditions log is usually what tells you whether the world changed or your setup did.
Tracking this is genuinely awkward, and pretending otherwise is how most reporting in this field goes wrong. There is no console. Answers vary between runs. Referral attribution is inconsistent between assistants. Anyone handing you a single confident number has hidden a great deal of variance behind it.
Influencing Sources You Do Not Own The highest value work sits on pages your team cannot edit. Review platforms, directories, forum threads and comparison articles carry disproportionate weight in generated answers, and getting represented accurately on them requires outreach, correction requests and occasionally patience with people who are not obliged to help.
A page asking how much something costs that says pricing depends on your requirements has answered nothing, and it will not be cited because there is nothing to cite. A range with the variables named is a real answer and gets quoted.
Second, the questions have to keep coming from customers rather than from the content calendar. Within a few months the temptation appears to invent questions to fill a schedule, and invented questions produce exactly the marketing-in-disguise sections that get ignored.
This variability is the main practical trap. Testing without web access and concluding you are invisible measures the training corpus rather than current retrieval, and the two can disagree sharply. Record which mode you used with every run.
Keeping It Honest Two disciplines keep this from decaying. First, the answers have to be checked by somebody who knows the business, because a writer working from notes will approximate a figure and ai seo services an approximation published as fact is a liability you carry rather than they do.
Report frequency rather than presence. Being named in one run out of five is a genuinely different situation from being named in five out of five, and a report that collapses both to mentioned has thrown away the useful part.
What analytics cannot tell you is how often you were named without a click, which in this channel is most of the time. A recommendation that a buyer acts on three weeks later leaves no trace in any report you own. This is why the manual prompt set is not optional, and why nobody should be asked to justify this work on referral traffic alone.
Why One Snapshot Proves Almost Nothing Generation involves randomness, and retrieval can return different pages between runs. The same prompt asked twice in a row can produce different companies in different orders.
One structural decision saves a lot of trouble later. Keep the raw answers in plain text files named by date, assistant and run number, rather than pasting them into a document that gets reformatted. Six months in you will want to search across every run for the first appearance of a competitor or a source, and a folder of plain files supports that while a slide deck does not.