Why ChatGPT Never Mentions Your Company
You Are Blocking the Crawlers The most common cause is also the least interesting. Your robots.txt disallows the user agents that feed AI systems, or a firewall rule is rejecting them, or a bot management product is serving them a challenge page they cannot pass.
The defensible position is to spend an hour on it if you like, and to spend the rest of the week on the things every system already reads: accessible pages, accurate Organization markup, consistent identity and content a machine can quote.
Treat markup as something with a maintenance cost rather than a one off implementation. Prices change, people leave, products are discontinued, and structured data quietly keeps asserting the old version long after the visible page has been updated. Adding a schema review to whatever process already updates your pages costs minutes and prevents the most damaging failure mode, which is confidently stating something that is no longer true.
Then add the structural markup, then check the whole thing with a reader in mind rather than a crawler. If a page has become harder for a person to use, something has gone wrong and the change should be reversed.
Repeat it quarterly. Identity work has slow feedback, because scattered mentions have to be re-crawled before they join up, and the temptation to abandon it after a month is strong. It is usually the change that unlocks everything else. generative engine optimization
Set Up So You Do Not Fool Yourself Open a signed out session, or a fresh one with memory and personalisation disabled. This matters more than anything else in the method. An account that has spent the week researching your own company will show you a flattering picture that has nothing to do with what a stranger sees.
Read the Source List Before the Prose Where citations are shown, list every domain and count how often each appears. This is the single most useful output of the whole exercise, and most people skip it because the prose is more interesting.
Expect the timeline to be uneven. Crawler access can change what an assistant sees within days, because retrieval happens at answer time. Identity consistency takes longer, since scattered mentions have to be re-crawled before they join up. Third party coverage is slowest of all and is the part you control least directly, which is exactly why it is worth starting on it before you need the result.
What needs you: factual accuracy. Somebody inside the business has to confirm the numbers, limits and claims before publication, because you carry the consequence of anything untrue being published about your own products.
Decide What the Result Means Four outcomes, each pointing somewhere different. Absent everywhere with a clean robots file and no third party listings usually means an identity and coverage problem. Absent with a blocked crawler or an empty non-JavaScript page means a mechanical problem, which is the good news outcome because it is cheap.
Bring one other person from the business, ideally from sales. They will spot inaccuracies in how you are described that a marketing reader skims past, and they will tell you within minutes whether the prompts sound like real customers. That second opinion costs half an hour and prevents the most common flaw in a self run audit, which is a set of questions written in the company's own language.
Days Fifteen to Thirty: Fix the Plumbing Someone technical checks that the crawlers feeding AI systems can reach your site, that your bot protection is not silently blocking them, and that your important pages contain real content without JavaScript running.
Test it rather than assuming. Load your key pages with JavaScript disabled and see what survives. If the product specifications, pricing, service areas and contact details vanish, that is what a machine reads.
Then ask it to name your leadership, your location and what you sell. Wrong answers here point at specific sources you can go and correct, which makes this one of the few diagnostics in the field that hands you a task list directly.
You can do this yourself in about half an hour, with no subscriptions and no technical knowledge. It will not be as thorough as a full engagement, and it is more than enough to establish whether you have a problem and roughly what kind.
Statistics without sources. This field circulates figures faster than it checks them, and a number arriving without a publisher, a sample size and a date should be discounted rather than repeated to your board.
But it is a claim, not evidence. Markup asserting that you own a profile only helps if that profile exists and points back. The pattern that works is reciprocal: your site names the profile, the profile names your site, and a third party source independently associates the two.
That matters most for the facts that establish identity, because those are the facts that let scattered mentions of you resolve into one record. It matters far less for content, where the model is going to read the prose anyway and is reasonably good at it.