AI search news ·

Do AI crawlers actually read llms.txt? DEJAN logged 4 big-three fetches against 1,150 for robots.txt in 30 days — our own ledger says 0 against 27

On the evidence of two independent server logs, almost never. DEJAN's Dan Petrovic published a census of every fetch of dejan.ai's three machine-readable files between 27 July and 15 August 2026: counting only Google, Anthropic and OpenAI crawlers, robots.txt was fetched 1,150 times, the site's OKF content bundle 23 times, and llms.txt 4 times — three by Googlebot, one by ClaudeBot, zero by any OpenAI agent. We opened our own crawler-hit ledger, live since 2026-08-17 01:48Z, and read the same three files over its first 39 hours: robots.txt 65 fetches by bot-shaped user agents, 27 of them from the same three vendors; llms.txt one fetch, from an agent that matches no recognised AI or search crawler; the sitemap 12. Google's John Mueller said it flatly in June 2025 — "FWIW no AI system currently uses llms.txt" — and fourteen months later two request logs, one large and one small, agree with him.

DEJAN's census, as published

The post (2026-08-15, Dan Petrovic; DEJAN is an AI-SEO consultancy, so this is a practitioner's own log, not a neutral lab — read it as such) counts fetches from dejan.ai's request log of three files: `/robots.txt`, `/llms.txt` (the llmstxt.org index for language models) and the site's OKF bundle at `/okf/`, its full machine-readable text. Method as stated: agents grouped by published User-Agent token; "Browser" is a normal browser UA, "Unrecognised" matches no known crawler or browser; the window is the 30 days ending 15 August, but logging began 27 July, so it holds 20 days of data; and the counts are fetches, "not renders, citations, or answer appearances." It was carried by Search Engine Roundtable's daily recap on August 17.

File (dejan.ai, 30 d to 2026-08-15)GoogleAnthropicOpenAIBig threeAll agents
robots.txt2436082991,15013,442
OKF bundle (/okf/)11111232,841
llms.txt3104157

Google is Googlebot plus GoogleOther and Google-CloudVertexBot; Anthropic is ClaudeBot (445 of the 608 robots.txt fetches), Claude-User (163) and anthropic-ai; OpenAI is OAI-SearchBot (295), GPTBot (1) and ChatGPT-User (3). The 157 llms.txt fetches from all agents came from 78 IP addresses and 13 agents: Browser 85 (54.1%), Unrecognised 40 (25.5%), curl 13 (8.3%), AhrefsBot 4, HeadlessChrome 4, Googlebot 3, Meta-ExternalAgent 2, and one each from ClaudeBot, MJ12bot, SemrushBot, Applebot, Bingbot and Amazonbot. Two details in the post are the ones to keep: the three Googlebot fetches of llms.txt "were followed by no page request from the same IP addresses," and OpenAI's agents — the ones that fetched robots.txt 299 times and the content bundle 11 times — never requested llms.txt once.

Our own ledger over its first 39 hours

Our crawler-hit ledger records every request whose user agent looks bot-shaped (matches `bot|crawl|spider|slurp|fetch|-user|…`) at the edge and classifies it against a self-declared-token table; browsers and curl are not recorded, so it is a crawler-only view and its window is short — it was born 2026-08-17 01:48Z and we read it at 16:44Z on the 18th. Our `/llms.txt` has been live since 2026-06-27 and is linked from nowhere on the site, in robots.txt or in the sitemap; a crawler finds it only by convention.

File (askedabout.com, 39 h to 2026-08-18 16:44Z)GoogleAnthropicOpenAIBig threeAll bot-shaped agents
robots.txt16 (Googlebot)1 (Claude-User)10 (OAI-SearchBot)2765 (+ PerplexityBot 1, bingbot 3, YandexBot 5, 29 unrecognised)
sitemap.xml4 (Googlebot)01 (GPTBot)512 (+ bingbot 3, 4 unrecognised)
llms.txt00001 (unrecognised bot-shaped agent, 2026-08-18 11:57Z)

Same shape at a hundredth of the scale: robots.txt is read continuously (OAI-SearchBot alone came for it ten times in 39 hours, Googlebot sixteen), the sitemap is read, and llms.txt is not — one fetch, from nothing we recognise. We will keep the count running; if any of ChatGPT-User, OAI-SearchBot, GPTBot, ClaudeBot, Claude-User or PerplexityBot fetches `/llms.txt` in the next 30 days, we will say so with the timestamp. The linked-from-nowhere caveat cuts both ways: it means our zero is partly a discovery zero — but robots.txt and the sitemap are also found by convention, and they were found.

What this means for businesses that care about AI visibility

Do ChatGPT, Claude or Google actually fetch llms.txt?

Rarely to never, on the two server logs we have. In DEJAN's 30-day census to 2026-08-15, Google, Anthropic and OpenAI crawlers fetched dejan.ai's llms.txt 4 times (Googlebot 3, ClaudeBot 1, OpenAI 0) against 1,150 fetches of robots.txt; on askedabout.com's ledger over its first 39 hours, those vendors fetched robots.txt 27 times and llms.txt zero times. Google's John Mueller wrote in June 2025 that no AI system currently uses llms.txt.

Who is fetching llms.txt files, then?

Mostly humans and tools. Of 157 llms.txt fetches from all agents on dejan.ai, 85 (54.1%) carried a browser user agent, 40 were unrecognised agents and 13 were curl; SEO crawlers (AhrefsBot, SemrushBot, MJ12bot) and one-off fetches from Applebot, Bingbot, Amazonbot and Meta-ExternalAgent made up most of the rest.

Should I still add an llms.txt file?

It costs nothing and does no harm, but neither log shows an AI vendor's agent using it, and the three Googlebot fetches DEJAN saw were followed by no page requests. Spend the hour on robots.txt allow rules for the AI agents you want and on pages that answer buyer questions in their first 200 characters — those are the things the logs and the citation data show being read.

See your number

See which businesses AI names when your client's buyers ask.

Running this for clients? The $249 agency 5-pack audits five businesses, white-labeled.

Check a client's AI visibility

Begin your check

Free · 60 sec

No account · No card · 3 buyer questions, 2 engines

By running a check you agree to our Terms and Privacy Policy.

Who runs this