There are two reasons you are absent from an AI answer. Either the crawler could not read your page, or the query never produced an AI answer at all. The first is fixable in a week. The second is not broken. Establish which one you have before spending anything.
This is usually not a ranking problem. Two unrelated failures produce the same symptom, and the repair for one does nothing for the other. Budgets get committed before anyone establishes which failure is in play.
Which of the two failures is yours?
- Failure one is supply. A crawler requested your page and received something it could not read, so none of your text entered the index that the answer is assembled from
- Failure two is demand. The query you tested does not produce an AI answer for anybody, so there is no answer to be missing from
| Supply failure | Demand failure | |
|---|---|---|
| Symptom | Absent from an answer that exists | No AI answer appears at all |
| Root cause | Page unreadable, blocked, or not indexed | Query does not trigger an AI surface |
| Where you check | Your own server and robots file | The live result for that exact query |
| Cost to fix | One developer afternoon | Nothing, it is not broken |
The order matters. Diagnosing supply when the real answer is demand is how a budget disappears into a problem that was never there.
What does a crawler actually get when it fetches your page?
We fetched the homepage of 30 Tamil Nadu engineering colleges as raw HTML with JavaScript disabled. Five came back without usable content. What counts as usable is our judgment, not a published standard.
| What the raw HTML contained | Homepages | Share |
|---|---|---|
| Nothing at all | 2 of 30 | 6.7% |
| Fragments only | 3 of 30 | 10.0% |
| No meta description | 12 of 30 | 40.0% |
| Description broken, generic, or another page's | 4 of 30 | 13.3% |
"Two homepages in thirty returned nothing an AI crawler could quote, and nobody viewing them in a browser would ever see it."
Google's JavaScript SEO documentation says why this still matters: server side rendering "is still a great idea because it makes your website faster for users and crawlers, and not all bots can run JavaScript".
Want this run on your own site? Here is what a free AI visibility audit covers.
See how you show up in ChatGPT, Gemini and Perplexity. Free AI visibility audit, your pages fetched the way a crawler fetches them.
Do the AI crawlers run JavaScript?
We read the current crawler documentation from three publishers on 10 August 2026 and recorded one thing: does the publisher state whether its crawler executes JavaScript?
- Google states that its renderer does, using an evergreen version of Chromium, and its generative AI guide adds that Google can process content within JavaScript as long as it is not blocked
- OpenAI documents four agents, their user agent strings and their published IP ranges, and says nothing about JavaScript execution
- Perplexity documents two agents, their IP ranges and firewall guidance, and also says nothing about JavaScript execution
Read OpenAI's crawler page yourself. You will find user agent strings and published IP ranges, which is what a network team needs, and no statement either way about rendering.
An absent statement is not proof that a crawler cannot render. It means there is nothing to rely on. If your text only exists after a script runs, your visibility rests on undocumented behaviour.
Are you blocked without knowing it?
Blocking GPTBot and blocking OAI-SearchBot are different decisions. OpenAI documents GPTBot as the agent for foundation model training and OAI-SearchBot as the agent for search, and states that sites opted out of OAI-SearchBot "will not be shown in ChatGPT search answers, though can still appear as navigational links".
Your firewall is the second place to look, and nobody in marketing has the login. Perplexity's own crawler documentation tells site owners they may need to explicitly whitelist its bots so they can reach the content.
The third is newer. Google's generative AI optimisation guide now states that "in addition to the technical requirements for Search, a site must be included in Search generative AI features in Search Console to be eligible for display in generative AI features on Google Search".
"Three separate switches, owned by three separate people, and not one of them reports a failure to anybody."
Is there an AI answer to be absent from?
Before repairing anything, confirm the answer exists. In Ahrefs SERP data for 160 Indian marketing queries, AI Overviews appeared on 74 of 80 informational searches. Of the 80 commercial searches in the same pull, 30 carried a city name, and none of those 30 returned an AI Overview.
"If your buyers search with a city name attached, absence from an AI Overview is the normal state of that result, not a defect in your site."
Record two things for every query you test:
- whether an AI answer appears at all, in the exact words a buyer would type
- whether it still appears once you add your city name to the same query
That segmentation, its method and its limits are set out in our analysis of which query types AI Overviews actually take.
What we would check, in order
- 1
Fetch your own homepage and two money pages as raw HTML with JavaScript disabled, and read what returns
- 2
Open robots.txt and confirm OAI-SearchBot and PerplexityBot are not disallowed alongside GPTBot
- 3
Ask whoever owns the firewall whether AI crawler user agents are being challenged or blocked
- 4
Confirm the site is included in Search generative AI features in Search Console
- 5
Only then test the queries, in the exact words a buyer would use, and record whether an AI answer appears at all
Steps one to four cost an afternoon and no money. Step five is the one that decides whether the other four were worth doing, which is also why we put technical checks ahead of content work on a first engagement.
What this cannot tell you
- The fetch covers 30 engineering college homepages in one Indian state. It is evidence about how that sector builds websites and nothing more. A retailer on a hosted platform will look different
- The SERP figures are Ahrefs' record of what its crawler saw, not our own searches, and they cover marketing services queries in India only
- We have not measured the thing that would settle the argument, which is whether repairing server side rendering changes how often a site gets cited. That needs a before and after on the same set of sites, and we have not run it
The full metadata breakdown from the same fetch sits in our teardown of Tamil Nadu college homepages. It records what was missing across all 30, in aggregate, with the named results held back.
Key takeaways
- Absence from an AI answer is two independent failures, and testing which one you have costs nothing
- Of 30 college homepages fetched with JavaScript off, 2 returned nothing at all and 3 returned only fragments
- Google publishes that its renderer executes JavaScript. OpenAI and Perplexity publish no such statement about theirs
- Blocking GPTBot is a training decision. OAI-SearchBot is the one that governs ChatGPT search visibility
- A firewall rule and a Search Console setting can both remove you silently, and neither reports a failure
Frequently asked
Does blocking GPTBot remove me from ChatGPT search?
No. OpenAI documents OAI-SearchBot as the agent used for search and GPTBot as the agent used for foundation model training, and treats the two settings as independent. If you want to appear in ChatGPT search answers, OAI-SearchBot is the one that has to be allowed.
Do I need an llms.txt file to appear in AI answers?
Not for Google. Its generative AI guide states you do not need new machine readable files, AI text files, markup or Markdown, because Google Search does not use them. Other systems may read such files. It is not the reason you are missing.
How do I check what an AI crawler sees on my page?
Request the raw HTML with JavaScript disabled and read what comes back. If your headline, your service list and your enquiry route are not in that text, they are not reliably available to anything that does not render.
How long after a robots.txt fix before anything changes?
OpenAI and Perplexity both document roughly 24 hours for a robots.txt change to register in their systems. Appearing in answers takes longer, because the page still has to be crawled, stored and then selected for a response.
Is being absent from an AI Overview always a problem?
No. In our sample of 160 queries, city qualified commercial searches carried an AI Overview zero times in 30. If your buyers search that way, absence is the normal state of that result rather than a fault in your website.
Does a missing meta description keep me out of AI answers?
We have no evidence that it does on its own. We report it because it travels with the same neglect. Of the 30 homepages we fetched, 12 carried no description and 4 carried one that was broken, generic or written for a different page.
Sources
Adith Krishnan
Co-Founder & COO, Kula Digital
8+ years building marketing that is measured in revenue. Runs strategy, AI search visibility, and education marketing at Kula. Written from the studio in Coimbatore.
