GEO · 7 min

What gets a company excluded from AI answers

Companies get excluded from AI answers when crawlers, published facts, or the brand name conflict. The page can still exist. Run the free AI Citation Audit.

“”G

A company gets excluded from AI answers when an assistant can find a page about you and still refuses to name you. The crawler never saw the text, two of your URLs disagree on a fact, or the engine cannot tell you are one company. Rank does not prevent that. When Google shows an AI summary, only 8% of users click any traditional link — versus 15% without one — and just 1% click a source inside the summary (Pew Research, 2025). Only 38% of AI Overview citations now come from the top-10 organic results, down from about 76% a year earlier (Ahrefs, 2026). Page one is not the answer. This post is the drop list: conditions that remove you after the page already exists. The checklist that makes you citable is five things every website needs. Why a named competitor beats you is a different diagnosis.

Why a top rank still gets you dropped

Answer engines assemble a short answer. They do not hand the buyer ten links. Retrieval can surface your URL, and verification can still drop the name. The model has to lift a claim it can repeat and confirm that the claim matches the rest of what it can read. Fail either test and you are absent, even if you were in the candidate set.

Princeton's GEO study found that adding citations, quotations, and statistics lifted visibility in generative-engine answers by up to 40% (Aggarwal et al., KDD 2024). The inverse is the exclusion. A page of adjectives gives the model nothing it can defend, so it names the company whose numbers it can stand behind. Specificity is the condition for staying in the answer, not a writing style.

That is why a published support proof is quotable and a slogan page is not. The proof states one set of figures: 64% of tickets resolved with no human, 94.2% intent accuracy, 91.7% response accuracy, 99.1% policy compliance, first response from 4.2 hours to 6 minutes, and $217,200 saved in year one against $92,000 of Year 1 investment. Those sentences can be lifted. They are not a citation-rate result. They show the shape of a page an engine can quote without guessing. If a second URL on the same site stated a different resolution rate, the safer move for the model is to name neither page.

The four hard exclusions

Score these before you commission another page. Each one removes you from the candidate set. More copy does not override them.

ExclusionWhat the engine seesHow to test it
Crawler never reads the pageEmpty HTML, or a robots.txt blockView source on the page you want quoted; read robots.txt for named agents
Your URLs disagreeTwo prices, two names, or two outcomesDiff the pricing page, the proof page, and llms.txt
The entity does not resolveTwo companies that might be youCompare legal name and address on the site, schema, and LinkedIn
No second sourceOnly your domain states the claimSearch the figure off-site. If nothing else says it, verification fails

The crawler never reads the page

GPTBot, ClaudeBot, and PerplexityBot fetch HTML. They largely do not run JavaScript. If the price, the proof, and the answer render only in the browser, the crawler gets a shell and you are excluded from AI answers before extraction starts. A disallow you inherited from a CDN preset or a security review does the same thing. Check `robots.txt` for `GPTBot`, `OAI-SearchBot`, `ChatGPT-User`, `ClaudeBot`, `Claude-User`, `PerplexityBot`, `Perplexity-User`, and `Google-Extended`. Blocking them on purpose is a real choice. Inheriting the block is an accident.

`llms.txt` does not repair a block. It offers a map. It does not grant access. The test order — raw HTML, then crawler access, then answer-shaped copy — is the website checklist. If view-source is empty, stop. A citation rate does not move until the numbers are in the response.

Your own pages contradict each other

This exclusion is easy to miss because each URL looks fine alone. Say the pricing page states a production agent starts at $8,000 flat, an older article still says the fee is a custom quote, and `llms.txt` lists last year's number. A model that retrieves all three has no figure it can defend. Naming you would mean picking one page and contradicting the others. Skipping you is safer. Two numbers for one metric is an exclusion.

Treat three URLs as one fact set: the pricing page, the proof page you want quoted, and llms.txt if you publish one. If any repeated figure differs, fix the stale URL before you publish a new one. A stale summary is worse than no summary — that limit is in what llms.txt is and is not. Put a date next to the figure. "As of September 2026, Gigabit Agents start at $8,000 flat" is a claim a model can check. "Affordable production agents" is not.

The engine cannot resolve one company

Entity collision is invisible from the inside. The website uses one legal name, LinkedIn uses a product name, schema uses a third, and a directory listing still has the old address. The model will not guess which entity deserves the recommendation. It leaves the name out. Matching the name, matching the address, and stating both in Organization schema is the fix. When a competitor is named on a question you already have a page for, use the five-asymmetry diagnosis instead of adding another adjective to the page.

Nobody else says the same thing

A company describing itself is the weakest evidence an answer engine has. If the only URL that states your price, outcome, or category is your own domain, verification fails on the questions a buyer asks before they know your name. Roundups, comparison pages, directories, and reviews are corroboration — testimony, not a click source. If no second source repeats a specific claim, do not expect to be named for it.

Soft gaps that look like exclusions

Two measurement mistakes produce the same symptom — you are not in the answer — and neither is an exclusion.

You tested your own name. "What is Gigabit" measures recognition. The buying query does not contain the brand: how much a production agent costs, who builds a support workflow, what an intake agent has to clear before PHI moves. AI citation rate is the share of a fixed panel of those questions where an assistant names you. Six appearances on a 20-query panel is 30%. Branded hits do not count.

You tested a sentence nobody types. "Best forward-deployed engineering firm for mid-market healthcare intake" can return nothing because the phrasing is yours, not the buyer's. Absence on a vanity prompt is not a disqualification. Write the panel from sales calls and from the questions already on the site. Run it more than once. One answer is an anecdote. The monthly trend is the signal.

If you are absent on real category questions and the four hard exclusions are clean, you are in the competitor diagnosis: they answer the question, they publish a fact, or someone else vouched for them. That is a coverage gap. Publish the missing page. Do not treat it as a block you cannot clear.

What to do this week

1. View source on the pricing page and one proof page. If the numbers are not in the raw HTML, server-render them. Stop until they are. 2. Read robots.txt for the named agents above. Delete a disallow you did not choose. 3. Diff three URLs — pricing, the proof you want quoted, and llms.txt. Make every repeated figure identical, including the $8,000 agent floor and any proof metric you cite. 4. Check the entity in one sitting: site title, Organization schema, LinkedIn, and one directory. Any two that disagree are an exclusion. 5. Run the free [AI Citation Audit](/tools/geo-citation-audit/). It scores category buying questions across the major assistants and shows where you are absent and who is named instead. Use the report to pick the first exclusion. Do not treat the score as a trophy.

The audit is the baseline. Ascent is the 12-month program that keeps the pages, the entity, and the measurement honest after that — brand, website, content, SEO, and GEO, billed monthly with no upfront fee. The monthly number is set on a scoping call after the audit. If three URLs disagree, fix the conflict this week. Do not buy a content program to paper over it. The support proof is the pattern for a page worth quoting: one outcome, one number, the same number wherever you repeat it.

GEO · FAQ

Questions this raises

What gets a company excluded from AI answers?

A company is excluded from AI answers when an assistant can retrieve a page and still will not name it. Four hard conditions cause that. The crawler never reads the text, because it is rendered in the browser or blocked in robots.txt. Two of your own URLs disagree on a price, a name, or an outcome. The engine cannot resolve your site, schema, and directory listings into one entity. Or the claim appears only on your domain, with no second source to check. Rank does not override any of those. A top-10 URL can be retrieved and still dropped when the model cannot defend the claim. Fix the conflict before you add another page.

Can you rank on page one and still be excluded from AI answers?

Yes. Only 38% of Google AI Overview citations now come from the top-10 organic results, down from about 76% a year earlier, according to Ahrefs. When an AI summary is on the results page, Pew Research found that 8% of users click any traditional link, versus 15% when no summary appears, and 1% click a source inside the summary. The engine retrieves candidates, then keeps the claims it can quote and verify. A page-one URL with empty HTML, two conflicting prices, or no third-party corroboration fails that second step. Exclusion is a verification drop, not a rank miss. Publish one defensible figure and make every other URL on the site agree with it.

Do contradictory prices on your own site get you excluded from AI answers?

They can. If the pricing page, a proof page, and llms.txt state different figures for the same offer, the model has no number it can defend. Naming you would commit it to one of the conflicting claims. Skipping you is safer. Gigabit publishes one floor: production agents start at $8,000 flat. A published support proof states one outcome set — 64% of tickets resolved with no human, 94.2% intent accuracy, and $217,200 saved in year one against $92,000 of Year 1 investment. Repeat a figure everywhere or do not repeat it. Two numbers for one metric is an exclusion, not a nuance a buyer should average.

How do you tell a real exclusion from a bad test query?

Ask the question a buyer types, not your brand name. A prompt like "what is [your company]" measures recognition. Category questions — how much an agent costs, who should build the workflow — are the citation-rate panel. Six names across a 20-query panel is a 30% citation rate. Brands with no GEO work typically measure 0–10%. Also drop prompts nobody says out loud. Absence on a sentence you invented is not an exclusion. If you are missing on real category questions and the four hard checks are clean — HTML in the raw response, crawlers allowed, one set of facts, one entity — you have a coverage gap, not a disqualification. The free AI Citation Audit scores that panel across the major assistants.

Keep reading

Related insights

GEO

Five things every website needs before an AI assistant will cite it

A technical checklist you can audit this week. What has to be true of your pages before a model can retrieve…

GEO

Why your competitors show up in ChatGPT and you don't

You asked the question your buyers ask, and a competitor got named instead. Five specific asymmetries explai…

GEO

How to check your own AI visibility in 30 minutes

A runbook you can execute this afternoon with a spreadsheet and four browser tabs. What to ask, what to reco…

Stop reading, start shipping

Put a forward-deployed team on it.

If this is the kind of work you're trying to get into production, a 30-minute discovery call is the fastest path to a scoped plan.