There is a question that shortcuts most of the confusion about AI visibility, and almost nobody asks it. Not how do I rank in ChatGPT — there is no ranking. The useful question is: when an assistant recommends a business in my category, where did that name come from?
Where an assistant shows its citations, this stops being theoretical. You can read the list.
The pages that keep appearing
Do this across enough categories and cities and the same shapes recur. Not the same websites — those differ by industry and country — but the same kinds of websites.
Directories and listings. The industry-specific ones matter more than the general ones. A trade association's member list, a professional register, a category-specific platform. These are dense with exactly the information an assistant needs: name, category, location, sometimes credentials, all in a consistent format across many businesses.
Roundups and "best of" articles. Someone wrote The 8 Best Physiotherapists in Bristol in 2023. It is possibly out of date, possibly written by someone who never visited any of them, and it is one of the most influential pages in that market, because its structure is precisely the structure of the answer the assistant is being asked to produce.
Review platforms. Not only for the star ratings, but because review pages carry a great deal of natural language about a specific business — the phrases customers actually use, attached to a name. That is unusually rich material for a model.
Forum and community threads. A discussion where someone asked your exact question and got five candid replies is, from a model's point of view, very high quality: real people, weighing real options, giving reasons.
Local press and institutional pages. Chamber of commerce membership, a local news piece, a supplier's "where to find us" page. Individually minor, collectively a strong signal that a business is real and locatable.
Why your own site ranks lower than you'd like
Owners consistently overestimate their website's role, because the website is the thing they control and the thing they paid for. But consider what an assistant is doing when it answers "which clinic should I use in Tijuana?" It needs several names with distinguishing reasons. A single business's website supplies one name and describes it in promotional language that fits every competitor equally well.
Your site is not useless. It is where an assistant confirms facts — what you do, where, for whom, at what price, with what qualifications. When it is vague or outdated, that confirmation fails, and a mention elsewhere never becomes a recommendation. But the site is rarely the reason you were named. It is the reason the naming survives scrutiny.
The practical consequence
Because the influential sources are a short list rather than the whole web, this becomes an addressable problem rather than a vague one.
For any given category and city, the citations tend to concentrate: a handful of sources drive most of the answers. Once you know which, you can work through them one at a time:
- Where you're listed but wrong. Extremely common, and the cheapest fix available. An old address, a service you stopped offering, a phone number from two premises ago. You already occupy the position; the information in it is just stale.
- Where you're absent but eligible. An association you belong to but never completed the profile for. A register that requires only an application.
- Where a competitor is and you aren't. A roundup with a named author who can be contacted. A review platform where you have a neglected profile.
- Where you cannot realistically appear. Some sources are closed, paid, or editorially hostile. Knowing this is still worth something: it tells you where not to spend effort.
That sequence is worth more than any amount of rewriting your homepage, and it is available to any business willing to look.
The catch
Citations are not the whole story. When an assistant answers from training data rather than live search, it cites nothing — and the influences are older, broader, and not directly visible. You cannot read them; you can only infer them from what the answer says.
This is why an answer with no citations should not be read as an answer with no sources. It usually means the sources are old enough to have been absorbed during training, which is both harder to change and more durable once changed.
It is also why a single check tells you very little. The same question asked twice can produce one cited answer and one uncited one, naming different businesses. Understanding which sources genuinely control your category takes repetition, across assistants, across phrasings — which is a measurement problem, not a browsing problem.