An AI assistant names a business when it can retrieve something about that business, verify which entity it belongs to, and find it relevant to the question asked. Failure at any one of those three stages produces the same symptom — a competitor gets named and you do not — which is why the problem is so often misdiagnosed.

Reason one: the assistant cannot reach your site

This is the one nobody checks, and it is invisible to every conventional audit.

On 20 September 2026 we tested eleven AI and search crawlers against seven domains. Ten received HTTP 200. One — a major AI system's corpus crawler, the one that builds its training and retrieval index — received HTTP 429 with an empty body on every single origin-bound request, while a normal browser received 200 on the identical URL in the same second.

The block was not in robots.txt, not in .htaccess, and not in any site code. It sat at the CDN edge, above the server, and the hosting provider confirmed it as a platform-side defect. The site owner had no way of knowing: the pages returned 200 to every browser and every other bot, and the provider's own dashboard reported the crawler as allowed.

Two things follow. robots.txt permission is not access — a crawler can be welcomed in robots.txt and refused at the edge. And testing robots.txt alone gives a false pass, because static cached files are often served to a crawler that is refused on everything else. The only honest test is to request an HTML page using the crawler's own user agent and read the status code.

Reason two: the assistant cannot tell which business you are

Assistants resolve entities before they recommend. If your business name, address and phone number disagree between your Google Business Profile, your website, your directory listings and your company filings, there is no single confident answer to "who is this", and a system with a low-confidence entity will reach for a higher-confidence one instead.

This is the most common cause and the most fixable. It is also the one that overlaps most with ordinary local SEO, which is why the work is not exotic: consistent details, a verified profile, structured data that describes the organisation and links to the places it is independently listed.

Trading names that differ from licensed or registered entity names are a frequent source of this, and in some markets — the UAE especially — it is close to universal.

Reason three: the question you care about cannot be won on your own page

In August 2026 we examined sixty records across thirty commercial queries and two engines, looking at what each engine actually cited. The finding was that the citation mechanism is predicted by query specificity, not query difficulty.

Specific queries tend to resolve to a vendor's own domain. Broad comparison queries — "best X company", "top X agencies" — resolve overwhelmingly to third-party listicles and roundups. For that second class, no amount of work on your own site changes the outcome, because your site is not the kind of source the engine is selecting.

So before investing in a page, it is worth establishing which class your question belongs to. Effort against a listicle-locked query is spent whether or not it works, and it will not work.

Reason four: nothing on the page is extractable

A page can be reachable, correctly attributed and on a winnable query, and still not be quoted, because there is nothing on it that survives being lifted out of context.

We looked at ninety days of AI referral data against page structure and position. Two findings held. Rank is not the gate — a page at position 70 took twenty-one referrals in ninety days while a near-identical page at position 69 took none. And structural markup is necessary but not sufficient: pages carrying the full feature set at higher impressions than the cited ones took no referrals at all.

That second finding matters because it cuts against how this is usually sold. Structure is worth building and it guarantees nothing. Anyone promising citations in exchange for schema is describing something they cannot control.

What to check, in order

Test crawler access properly. Request an HTML page with the crawler's user agent and read the status code. Not robots.txt, not a settings panel — a real page.

Resolve the entity. Make the name, address and phone identical everywhere, verify the profile, and make sure the structured data on your site says who you are and links to where you are independently listed.

Classify the query. Establish whether the question you want to win resolves to vendor pages or to listicles, before building anything for it.

Then make the answer extractable. A self-contained, accurate sentence that is true standing alone is the unit that gets quoted. A sentence that needs the paragraph around it to be accurate is not a citation candidate, and compressing it into one anyway is how businesses end up with a machine-readable claim they cannot stand behind.

What this is not

It is not a guarantee of being recommended. Nobody controls what an assistant names, and any agency saying otherwise is selling something it cannot deliver.

It is not a replacement for technical SEO, local search work, paid media or conversion work. It is the same disciplines applied with retrieval in mind.

And it is not a reason to rename ordinary SEO as "AI SEO". Most of what makes a business retrievable by an assistant is work that was worth doing anyway.

Quick Answers

Why does ChatGPT mention competitors and not my business?

Four distinct causes: the assistant cannot reach your site, cannot tell which business you are, the query resolves to listicles rather than vendor pages, or nothing on the page is extractable.

Does robots.txt permission mean an AI crawler can reach my site?

No. A crawler can be permitted in robots.txt and refused at the CDN edge. Only requesting a page with that crawler's user agent tests access.

Can an AI crawler be blocked without the site owner knowing?

Yes. A block at the CDN edge appears in no site configuration and returns no error to the site owner.

Will schema markup alone get a business cited?

No. Pages carrying the full structural feature set at higher impressions than cited pages took no referrals at all.

Are broad 'best company' queries winnable on your own site?

Generally not. They resolve overwhelmingly to third-party roundups rather than to any vendor's own pages.

Frequently Asked Questions

There are four distinct causes and they need different fixes: the assistant cannot reach your site, it cannot confidently tell which business you are, the question resolves to third-party listicles rather than vendor pages, or nothing on your page is extractable as a standalone statement.

Yes. In testing across seven domains we found a major AI corpus crawler receiving HTTP 429 on every origin-bound request while ten other crawlers received 200 on identical URLs. The block sat at the CDN edge, above the server, so it appeared in no site configuration and the hosting dashboard reported the crawler as allowed.

No. robots.txt permission is not access. A crawler can be permitted in robots.txt and refused at the CDN or WAF layer. The only reliable test is requesting a page with that crawler's user agent and reading the status code.

Static files are often served from cache to a crawler that is refused on everything else, so robots.txt can return 200 while every HTML page returns an error to the same crawler.

Not on its own. In ninety days of referral data, pages carrying the full structural feature set at higher impressions than the cited pages took no referrals at all. Structure appears necessary and is demonstrably not sufficient.

Not reliably. A page at position 70 took twenty-one AI referrals in ninety days while a near-identical page at position 69 took none. Rank and citation are related but not the same thing.

Broad comparison queries such as 'best X company' resolve overwhelmingly to third-party roundups rather than to any vendor's own pages. For that class, work on your own site does not change the outcome.

Entity ambiguity. When your name, address and phone disagree across your profile, your site and third-party listings, no system can confidently say who you are, and it will name a business it is more certain about.

Talk to someone

Call (844) 677-1981 or email sales@rankure.io. Rankure LLC, 1550 Wilson Blvd, Arlington, VA 22209.