Skip to content
SEO Madmanby Adam Hafez

Mueller: AI crawlers read sitemaps, Couldn't fetch explained

Google's John Mueller says AI crawlers can't submit sitemaps and that Couldn't fetch on a valid sitemap often comes down to host load or crawl demand.

Published · 3 min read

Written byAdam Hafez
A wooden signpost with three directional trail signs above a coastal hillside

Key takeaways

  • Mueller said AI training crawlers usually have no console to submit a sitemap, so sites that want them to find content should keep sitemap.xml as the name or publish RSS feeds.
  • Mueller said he has seen an AI crawler fetch his sitemap and RSS files in his server logs, but he did not name the crawlers or say what they do with the files.
  • A private sitemap with an unusual name, left out of robots.txt and submitted to Google directly stays hidden from other systems, and Bing would need its own submission.
  • Search Console can report Couldn't fetch for a valid, public sitemap, and Mueller gave two causes, which are Google's host load and crawl demand tied to perceived site quality.
  • Search Engine Journal notes both causes sit outside the sitemap file, so a clean validation does not rule the error out.

Google’s John Mueller and Martin Splitt spent the 1 October 2026 episode of Search Off the Record, Do sitemaps still matter?, on two sitemap questions. Search Engine Journal reported what Mueller said in one article by Matt G. Southern and one by Roger Montti. We did not hear the episode: its YouTube page confirms the title and the 1 October upload, and every statement below is as reported by those two writers.

Can AI crawlers find your sitemap?

Per Southern, Mueller said AI training crawlers usually have no console or setup where a site owner can submit a sitemap file. For sites that want their content in AI systems, he suggested either keeping the generic name sitemap.xml or focusing on RSS feeds. Feeds are easier to find, he noted, because pages usually link them from the HTML head.

He also said he has seen an AI crawler access his sitemap in his server logs, and the same with his RSS files. He did not name the crawlers, and said he does not know whether AI companies document this or what they do with the files.

Southern adds that the sitemaps protocol treats the robots.txt Sitemap line as independent of the user-agent line, so it is not tied to one crawler’s rules. Check the pair with the sitemap vs robots.txt checker.

What about a private sitemap?

Mueller said a site owner who wants a sitemap kept private can give it an unusual file name, leave it out of robots.txt and submit it to Google directly. The downside, per Southern’s report, is that other systems cannot find it, and Bing would probably need its own submission.

Does llms.txt replace a sitemap?

Asked that, Mueller compared the Markdown file to an HTML sitemap and said Google’s systems cannot use it as a sitemap because it lacks the strict format. Southern reports he said the hope is bigger than the reality, that none of this happens today, and that he would not rely on it. See the W3C llms.txt draft for where the format stands.

Why do valid sitemaps fail to fetch?

Splitt put the question: the XML validates, it is public and robots.txt links it, yet Search Console sometimes says it can’t fetch it. Per Montti, Splitt acknowledged this happens to many people with valid sitemaps. Mueller gave two reasons:

  • Host load: Google’s systems may be too busy with other crawling to fetch the sitemap, and Search Console reports that as Couldn’t fetch too.
  • Crawl demand: if Google’s systems see no need to crawl much more from a site, they may skip the sitemap. Mueller said crawl demand is very often based on perceived quality, so it is not purely a technical thing.

Montti reads a third point into Mueller’s answer: Google may simply not need the sitemap, and would use it if site quality improved significantly. Montti’s own view is that the message does not match the actual reason, which leaves Search Console users confused, and that Google knows this and has done nothing. That is his opinion, not Google’s.

What should you rule out first?

Google’s Sitemaps report help page lists low crawl demand among the reasons a fetch fails, and says the higher the quality of the site’s content, the higher the crawl demand. It also lists a robots.txt block, an unresolved manual action and a wrong URL that returns a 404.

Those are the checks you can run: confirm the file parses with the XML sitemap validator, then look in your server logs for failed requests around the fetch time. Southern points to logs as where Mueller saw AI crawlers, and where you can check your own. If everything is clean, the reporting points to demand and quality, not the file.

The evidence

Type
industry
Impact
medium
Affects
XML sitemaps, RSS feeds, AI crawlers, Search Console sitemap errors

Sources

  1. 1.Do sitemaps still matter? - Google Search Central (YouTube), October 1, 2026Primary
  2. 2.Google's Mueller Says AI Crawlers Access Sitemaps & RSS In His Logs - Search Engine Journal, October 5, 2026
  3. 3.Google Explains Couldn't Fetch Sitemap Errors - Search Engine Journal, October 6, 2026
  4. 4.Sitemaps report - Search Console Help - Google Search Console Help

Frequently asked questions

Can AI crawlers find my sitemap?

Possibly, per Search Engine Journal. John Mueller said AI training crawlers usually offer no way to submit one, so keep the generic sitemap.xml name or publish RSS feeds. He has seen an AI crawler fetch his sitemap and RSS files in his logs.

Why does Search Console say Couldn't fetch for a valid sitemap?

Mueller gave two reasons, as reported by Search Engine Journal. Google's systems may be too busy to fetch the file, or crawl demand is low, which he said is very often based on how Google perceives the site's quality.

Can I keep a sitemap private?

Per Mueller, yes. Give it an unusual file name, leave it out of robots.txt and submit it to Google directly. The cost is that other systems can't find it, and Bing would probably need its own submission.

Does llms.txt work as a sitemap?

Not for Google, according to Mueller as reported by Search Engine Journal. He compared it to an HTML sitemap and said Google's systems can't use it as a sitemap because it lacks the strict format.

About the author

Adam Hafez
Adam Hafez

Founder, UpgradIQ FZC LLC

Adam Hafez works on technical SEO and search measurement: how pages get crawled, indexed, ranked and now quoted by answer engines. He founded UpgradIQ, which reads Google Search Console and GA4 to tie ranking movement back to the changes that caused it. He publishes what the data supports and states the limits of it.

  • Technical SEO
  • Search Console and GA4 measurement
  • Answer engine optimization
  • Structured data

The briefing

Only what actually changed in search, delivered in full by RSS, Atom or JSON feed.

Follow
Tools

Hreflang XML sitemap generator

Hreflang sitemap generator: turn language codes and URLs into sitemap XML with xhtml:link alternates, and check codes and URLs. Free, runs in your browser.