Service · AI search
AI search optimization, from eligibility to control
AI search optimization is the technical groundwork under GEO and AEO: making sure AI engines can find, read and quote your pages, measuring when they do, and making deliberate choices about the controls Google now offers. For lawful adult brands, it also means keeping quotable pages safe for work.
- 01Eligible
Indexed, crawlable, snippet-eligible, safe for work where it should be quoted.
- 02Reachable
The right AI search crawlers allowed; training crawlers decided on separately.
- 03Measured
The new Search Console AI report, plus checks in ChatGPT, Perplexity and Gemini.
- 04Controlled
Deliberate choices on the AI opt-out and snippet controls, not accidents.
Eligibility comes first
For Google, a page can support an AI Overview or AI Mode answer only if it is indexed and eligible for a snippet; there are no extra technical requirements. For other engines, the right search crawlers must be allowed: OAI-SearchBot for ChatGPT search and PerplexityBot for Perplexity, for example.
Adult brands have one more check. Pages that SafeSearch treats as explicit are filtered for people who use it, so the pages you want quoted should be clearly safe for work, with explicit material labeled and kept elsewhere. Crawler choices are laid out on adult GEO.
What changed in 2026
- 3 June 2026
Search Console begins showing a generative AI performance report and testing an AI opt-out, the same day a UK regulator requires one.
- 31 August 2026
Google confirms the site-level "Search generative AI" setting has rolled out to all websites worldwide.
- March 2027
The UK regulator's deadline for Google to add page-level controls.
It works on the whole site at once, and Google states that choosing to exclude a site has no effect on how it ranks in regular results. The report shows impressions by page, country, device and date, without clicks or queries. Snippet-level controls are compared on adult AEO.
About llms.txt
An llms.txt file is a plain-text map of a site's key pages written for AI tools. It is easy to add, but Google's guidance for AI features says no new machine-readable AI files are needed to appear in them.
We add one when a client wants it, and we do not sell it as a ranking fix. The effort goes into pages that can be indexed and quoted first. Measuring the effect across engines is covered on AI citation optimization.
Three kinds of AI crawler
The major AI companies now separate their bots by job: one collects training data, one builds the index their search features cite, and one fetches a page when a user asks the assistant to read it. Each needs its own decision, and engines differ in which sources they lean on, as our note on how AI engines cite adult brands shows.
| Company | Training | Search index | User-requested fetch |
|---|---|---|---|
| OpenAI | GPTBot | OAI-SearchBot | ChatGPT-User |
| Anthropic | ClaudeBot | Claude-SearchBot | Claude-User |
| Perplexity | No separate bot listed | PerplexityBot | Perplexity-User |
| Google-Extended token | Googlebot | Google's own fetchers |
The user-requested column behaves differently. OpenAI's documentation says robots.txt rules may not apply to ChatGPT-User, because a person started the visit; Anthropic's documentation says all three of its bots follow robots.txt; Perplexity treats its user fetches as user-initiated too. If something must never be read by any assistant, robots.txt is the wrong tool: it has to sit behind a login. Bot names and rules change, so we recheck this table against each company's documentation every quarter.
A robots.txt that matches your choices
Most lawful adult brands we work with want to be found in AI search and are less sure about training. A common pattern allows the search bots, makes a deliberate choice on the training bots, and keeps every bot away from account, checkout and member areas:
robots.txt, illustrative
User-agent: OAI-SearchBotUser-agent: Claude-SearchBotUser-agent: PerplexityBotAllow: /Disallow: /account/Disallow: /checkout/Disallow: /members/
Training bots such as GPTBot and ClaudeBot get their own group below it, allowed or disallowed as you decide. Two cautions. Blocking Googlebot to keep out of AI Overviews also removes you from Google Search; the Search Console setting exists for that choice. And a broad rule written years ago against all unknown bots may now be blocking the search crawlers you want.
Checking that the rules actually work
- Read your server logs for the bots above, and confirm which ones visit and which pages they request.
- Check that a visitor claiming to be a bot really is one: OpenAI publishes the IP ranges its crawlers use, and Google documents how to verify Googlebot.
- Look at your CDN and firewall. Bot protection settings can block AI crawlers before robots.txt is ever read, and some providers now offer AI crawler blocking as a simple switch.
- Ask each assistant to summarize one of your public pages and see whether it can; the full list of checks is in our AI search readiness checklist.
In our audits, a forgotten firewall rule is a more common reason for being invisible to AI search than anything in robots.txt.
Content that needs JavaScript to appear
Googlebot renders JavaScript; many AI crawlers do not. An analysis of crawler behavior published by Vercel and MERJ in late 2024 found that the major AI crawlers it studied fetched JavaScript files but did not execute them. If your product details, prices or answers only appear after a script runs, those crawlers may see an empty page. Serving the key text in the initial HTML, through server-side rendering or static generation, is the safer design, and it helps speed and accessibility too. Structured facts about your brand belong there as well, as covered in entity optimization.
When crawlers arrive faster than your server likes
AI crawlers can request pages in bursts, and on a small store or membership site that can slow things down for real visitors. Blocking them outright solves the load problem by creating a visibility problem. There are gentler tools. Anthropic documents that its bots respect the non-standard Crawl-delay rule in robots.txt; Google ignores that rule and adjusts its own pace when a server slows or returns errors. Caching public pages, serving them from a CDN and returning a proper 429 or 503 status under strain, rather than an error page with a 200 status, all help crawlers back off without forgetting you.
We look at crawler load in the logs before recommending any block, and when a block is right we make it as narrow as possible: one bot, one section, for as long as needed. If you want us to look at your own logs, request a readiness audit.
How progress is judged
Each month we report which AI crawlers visited and how often, any that were blocked by mistake, how many of your key pages return their main content without JavaScript, impressions in Google's AI features from the Search Console report, and whether the main assistants can read and summarize your important pages. How those pages then perform as sources is tracked on AI citation optimization; plans that include this monitoring are listed under pricing.
Questions
Q.01
Should we add an llms.txt file?
It is cheap and harmless, but do not expect it to change Google results. Google's guidance for AI features says no new machine-readable AI files are needed. Spend the effort on crawlable, indexable pages first.
Q.02
Should adult brands opt out of Google AI features?
Usually not. For a lawful adult brand, AI answers are a new way to be found. Opting out removes the whole site from those features, so use the performance report before deciding.
Q.03
Does SafeSearch affect AI answers?
Explicit pages are filtered for people with SafeSearch on, and AI features draw on the same Search systems. Keeping the pages you want quoted safe for work is the simplest protection.
Q.04
What do the new Search Console reports show?
Impressions in AI features by page, country, device and date. They do not show clicks or queries, so pair them with your analytics.
Q.05
Will blocking GPTBot remove us from ChatGPT search?
No. GPTBot collects training data; ChatGPT search relies on OAI-SearchBot. You can block one and allow the other, and OpenAI documents them separately.
Next experiment
What do AI answers say about your brand today?
- Hypothesis
- AI engines describe you less accurately than your site could support.
- Method
- We ask the questions your customers ask, across the main AI engines.
- Result
- A short readiness report with what to fix first. Free.