AI search news ·
Three of three crawler readings we pre-registered on 2026-08-26 held in week 3: ChatGPT-User 71% homepage (bar 50%), Googlebot 41% robots and sitemap (bar 40%), Amazonbot 166 distinct content pages against the AI fetchers' best of 69
All three readings we pre-registered on 2026-08-26 held on the week-3 ledger (2026-08-26 13:37 to 2026-09-02 13:30 UTC, 2278 net bot requests): ChatGPT-User 71% homepage against a 50% bar, Googlebot 41% robots.txt and sitemap.xml against a 40% bar, and Amazonbot 166 distinct content pages against 69 for the widest AI fetcher. Googlebot's plumbing share sits 0.5 points above its bar, the narrowest margin of the three, and it fell each week (56%, 46%, 41%). Method: our first-party crawler ledger, the same script as weeks 1 and 2; every number here is generated from the three committed artifacts (crawler-allocation-cut-2026-09-02.json, 9 of 9 self-checks).
What was pre-registered, and what the ledger says
The week-1 post (2026-08-26) closed with three readings to be scored on 2026-09-02 from a fresh 7-day dump, before the data existed. (A) ChatGPT-User sends at least half of its requests to the homepage; falsified below 30%. (B) Googlebot spends at least 40% of its requests on robots.txt and sitemap.xml. (C) Amazonbot touches at least as many distinct content pages as any single AI fetcher. The dump is 2026-08-26 13:37 to 2026-09-02 13:30 UTC: 2363 rows, 85 verifier probes excluded, 0 spoofed, 2278 net requests over 176 distinct paths. It does not overlap week 1; it overlaps week 2 (which ran 08-23 to 08-30) by four days, stated here because the series is a rolling read.
| Reading | Bar | Week 3 | Result |
|---|---|---|---|
| (A) ChatGPT-User homepage share | ≥ 50% holds; < 30% falsifies | 219 of 309 requests, 71% | HELD |
| (B) Googlebot robots.txt + sitemap.xml share | ≥ 40% | 109 of 269 requests, 41% (robots.txt 92, sitemap.xml 17) | HELD |
| (C) Amazonbot distinct content pages vs each AI fetcher | Amazonbot ≥ every AI fetcher | 166 vs ChatGPT-User 40, OAI-SearchBot 67, PerplexityBot 69, ClaudeBot 15, Claude-User 4, GPTBot 0, YouBot 32 | HELD |
3 of 3 held. Reading (B) is the one to watch: 41% clears a 40% bar by 0.5 points, down from 56% in week 1 and 46% in week 2. Googlebot made more requests each week (176, 240, 269) and put more of them on content (71, 121, 152 content requests over 45, 55, 56 distinct pages). On this trend the reading fails in week 4 or 5, which is the point of writing the bar down first.
Method, unchanged from week 1
The site's edge middleware forwards every bot-shaped User-Agent to a Convex table with the path, the token that matched, and a timestamp; no IP is stored in the artifact. A fixed token table names the fetcher by substring, so identity is claimed by the User-Agent string, not proven. A request is one HTTP request; nothing is de-duplicated. Path classes are assigned by a fixed regex in the script: homepage is `/`; plumbing is robots.txt, sitemap files, llms files and the well-known set; content is every `/news`, `/compare`, `/guide`, `/data`, `/for`, `/library` and site page. "Distinct content pages" counts unique paths in the content classes. The script is `crawler_allocation_cut.mjs`, the same file for all three weeks; the week-3 artifact passes 9 of 9 self-checks (partition sums, per-fetcher class sums, the 7-day span, and unit cases for the path and User-Agent classifiers).
Three weeks, eleven fetchers
| Fetcher | Requests w1 | w2 | w3 | Homepage w1 | w2 | w3 | Plumbing w1 | w2 | w3 | Distinct content pages w1 | w2 | w3 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|
| ChatGPT-User | 272 | 274 | 309 | 71% | 74% | 71% | 0% | 0% | 0% | 34 | 36 | 40 |
| Googlebot | 176 | 240 | 269 | 3% | 3% | 3% | 56% | 46% | 41% | 45 | 55 | 56 |
| YandexBot | 55 | 116 | 202 | 18% | 22% | 18% | 45% | 57% | 68% | 18 | 22 | 23 |
| Amazonbot | 164 | 168 | 177 | 1% | 2% | 3% | 0% | 0% | 0% | 151 | 158 | 166 |
| ClaudeBot | 139 | 250 | 169 | 1% | 0% | 0% | 46% | 67% | 91% | 66 | 73 | 15 |
| OAI-SearchBot | 129 | 165 | 162 | 3% | 3% | 2% | 36% | 32% | 36% | 57 | 69 | 67 |
| bingbot | 124 | 121 | 109 | 3% | 2% | 1% | 23% | 26% | 30% | 48 | 45 | 39 |
| PerplexityBot | 126 | 102 | 104 | 3% | 4% | 3% | 9% | 6% | 6% | 67 | 62 | 69 |
| YouBot | 0 | 59 | 59 | 0% | 2% | 2% | 0% | 44% | 44% | 0 | 32 | 32 |
| Claude-User | 25 | 23 | 15 | 0% | 0% | 0% | 44% | 39% | 33% | 5 | 4 | 4 |
| GPTBot | 8 | 8 | 8 | 13% | 13% | 13% | 88% | 88% | 88% | 0 | 0 | 0 |
Windows: week 1 2026-08-19 13:53 to 2026-08-26 13:37 UTC (1955 net), week 2 2026-08-23 14:06 to 2026-08-30 13:39 UTC (2146 net), week 3 2026-08-26 13:37 to 2026-09-02 13:30 UTC (2278 net). Named fetchers were 1227, 1532 and 1612 of those; the rest are bot-shaped User-Agents matching no token (PetalBot, SemrushBot and AhrefsBot lead that bucket every week).
Three movements. ClaudeBot went from 139 requests at 46% plumbing to 169 at 91%: it now polls robots.txt and the sitemap and reads 15 pages a week, down from 66. YandexBot went from 55 to 202 requests, 68% of them plumbing. ChatGPT-User grew from 272 to 309 requests with the homepage share flat at 71%, 74%, 71%; it touched 34, 36, 40 distinct content pages. GPTBot is unchanged at 8 requests a week, 7 of them robots.txt, 0 content pages in any week.
Pre-registered re-read, 2026-09-09
Same script, the ledger window 2026-09-02 13:30 UTC to 2026-09-09, scored in the growth run that day. Three readings, written before the data exists.
- 1(A) unchanged: ChatGPT-User homepage share at least 50%; falsified below 30%.
- 2(B) tightened to the trend: Googlebot plumbing share stays at or above 40%. The stated expectation is that it does NOT: the series is 56%, 46%, 41%, and a fourth reading below 40% ends the "Googlebot spends most of its budget on plumbing" line in this series. A reading at or above 40% keeps it.
- 3(C) unchanged: Amazonbot distinct content pages at least equal to every single AI fetcher's. The nearest AI fetcher this week is at 69; a fetcher crossing 166 would fail it.
Limits
- The ledger sees requests that reach the edge middleware. Cached responses never reach it, so every count is a floor.
- Fetchers are named by User-Agent, not reverse DNS. The known spoof-path scanner shape is excluded and counted; this week it was 0 rows.
- Week 2 and week 3 overlap by four days; week 1 and week 3 do not overlap. The three-week table is a rolling read, not three independent samples.
- The bars were set on week 1 alone, and (B) was set 16 points under the week-1 value. A bar that holds because it was set loosely is a weaker result than one set at the margin; the 09-09 read states the expected failure for (B) so the next result carries information either way.
- Artifacts: `content/news/crawler-allocation-cut-2026-09-02.json` (9 of 9 self-checks) from `crawler-rows-slim-2026-08-26-to-09-02.json`; the week-1 and week-2 artifacts are `crawler-allocation-cut-2026-08-26.json` and `crawler-allocation-cut-2026-08-30.json`. This post's text and tables were generated from the three artifacts by `crawler_allocation_week3_post.mjs`.
Did the three pre-registered crawler readings hold?
Yes, 3 of 3 on the week-3 ledger (2026-08-26 13:37 to 2026-09-02 13:30 UTC). ChatGPT-User 71% homepage, Googlebot 41% robots and sitemap, Amazonbot 166 distinct content pages against 69 for the widest AI fetcher.
Which reading is closest to failing?
Googlebot's plumbing share, at 41% against a 40% bar, down from 56% and 46%. The 2026-09-09 read states that it is expected to fall below 40%.
Does ChatGPT-User read the site or just the homepage?
71% of its 309 requests this week were the homepage; it touched 40 distinct content pages, 80 requests on /news posts.
Which crawler reads the most pages?
Amazonbot, at 166 distinct content pages in 7 days (177 requests, 97% content). The widest AI fetcher, PerplexityBot, reached 69.
See your number
See which businesses AI names when your client's buyers ask.
Running this for clients? The $249 agency 5-pack audits five businesses, white-labeled.
Who runs this
- Built and operated by Sensara LLC, Atlanta, Georgia — about us and how the audit works.
- See what the report looks like before you run anything — score per engine, the competitors AI names instead of you, and a fix plan.
- We run the same audit on ourselves every week and publish the result: in the latest run AI named AskedAbout in 1 of 144 answers. We report our own numbers the way we report yours.