AI search and WordPress
What is known about being found and quoted by AI assistants and by Google's AI answers, kept apart from what is only said. Start with whether their crawlers can reach your site, then what each engine documents.
- By
- WP Ministry
- Published
More and more questions are answered by an assistant or by a summary at the top of a search page, with a few links to where the answer came from. This subject is about being one of those links, for a site that runs on WordPress.
A great deal is said about it, and little of it by the engines themselves. So every page here says what each statement stands on: something an engine documents, something a study found, or guesswork. Google's own position is the plainest. It says that getting into its AI answers is the same work as getting into search, and that nothing extra is required.
Nothing here promises that an assistant will name your site. Nobody can promise that.
Start here. Generative engine optimization: what is documented, what was studied and what is guesswork is the whole picture in one place.
Google's AI answers. How to appear in Google's AI Overviews covers what Google says is needed, the report in Search Console that shows where you appeared, and the switch that takes a site out.
Their crawlers. An assistant cannot quote a page its crawler cannot reach. Blocking or allowing AI crawlers on WordPress covers robots.txt, a CDN or a host that blocks them before WordPress is asked, and how to tell which you have. The file itself is explained in where robots.txt is in WordPress.
llms.txt. A plugin may already have made you one. llms.txt on WordPress says what it is, who reads it and whether to bother.
The basics still decide it. A page that is blocked or asks not to be listed is out of AI answers too. Check with the page indexing check, and see a WordPress site that is not on Google.
Guides
- AI bots slowing down your website: how to confirm it on WordPress and what to do, mildest firstCount the requests in the server's access log by user agent. If AI crawlers lead, ask them in robots.txt to stay out of the addresses that cost the most, answer named crawlers with 429 while the server is overloaded, and block one outright only if it keeps coming.
- Bing Webmaster Tools for WordPress: setup, IndexNow, and what it means for ChatGPT and CopilotMicrosoft documents that Copilot rests on Bing's index. OpenAI documents its own crawler and does not say a site must be in Bing. Bing Webmaster Tools is free, a WordPress site is verified with one tag or one file, and its AI report shows phrases that Google's does not.
- Generative engine optimization (GEO): what is documented, what was studied and what is guessworkGenerative engine optimization (GEO, also sold as AEO or AI SEO) is the work of getting a site quoted in AI answers. What the engines document is short. Let their search crawlers in, keep pages indexable, put answers in visible text. The rest is a study or guesswork.
- How to appear in Google AI Overviews: what Google requires and what keeps a WordPress page outGoogle says a page needs nothing extra to appear in AI Overviews or AI Mode. It has to be indexed and allowed a snippet. On WordPress the work is checking that nothing takes a page out, such as a noindex, a snippet rule or the Search generative AI control in Search Console.
- How to block AI crawlers on WordPress, or let them in: robots.txt, Cloudflare, hosts and pluginsAI companies send different crawlers for training, for search and for fetching a page someone asked about. A WordPress site can refuse one kind and admit the others in robots.txt. A CDN, a host or a plugin can also block them where robots.txt does not show it.
- How to stop AI training on your WordPress content and stay in AI search answersYou can ask the crawlers that collect for AI training to stay out, one company at a time, and stay in most assistants' search answers. The request is a robots.txt group or, at Microsoft and Amazon, a noarchive tag. It is not a lock, and it does not take back what was already collected.
- llms.txt on WordPress: what it is, who reads it and whether your site needs onellms.txt is a proposed markdown file at /llms.txt that lists a site's main pages for AI agents. It is not a standard. Google says Search ignores it, and no AI company's crawler page says its crawler reads yours. An SEO plugin may already have made one, so check first.
Tools
- Are AI crawlers allowed on my site? A robots.txt checkGive your site's address and see what its robots.txt tells each AI crawler by name: the ones behind ChatGPT's, Claude's and Perplexity's search, the fetches made for their users, and the ones that collect pages for training.
- robots.txt tester: does this rule block this address?Paste a robots.txt, give an address and choose a crawler, and see whether it is allowed and which line decides. Worked out in your browser; nothing is sent.
- robots.txt generator for WordPressBuild a robots.txt that starts from WordPress's own rules, names your sitemap, and keeps out the kinds of AI crawler you choose. Made in your browser; nothing is sent.

