Skip to main content
LovedByAIThe Lab
Live
11 Sep 19:09 UTC
ChatGPT-User · 11sPerplexityBot · 1mOAI-SearchBot · 1mClaude-User · 3mDuckAssistBot · 19sGrokBot · 10m
Open research · AI search

Everyone asks the models. We change the page and read the server logs.

A small business should turn up when someone needs it. We measure what actually gets one found, from request-level logs on real small and medium-sized sites.

Someone asked · an assistant read the page+15.6% a month

Median user-agent fetches per site, per month

JanFebMarAprMayJunJulAug
January 2026 – August 2026 · fitted +15.6%/mo, R² 0.809Aug: 75 · 91.8% of sites reached
Assistant visits to a page · since January
2.76×more often

The median site went from 27 to 75 user-triggered fetches a month. Fitted growth is 15.6% a month at R² 0.809, which is a real trend rather than two lucky endpoints.

Who is reading · August 2026
96.9%OpenAI crawls

OpenAI leads by a distance: its crawlers reached 96.9% of monitored sites in August 2026, and on the middle site they fetched 132 pages that month. No other assistant comes close.

Turned away without knowing · 314 sites
35.7%of sites

block an AI crawler without any sign of it. Their robots.txt invites the assistant in and the server then refuses it, so the owner has no way to tell. 43% refuse at least one crawler, and only 5 owners wrote a rule themselves, and most of the written rules came from Cloudflare’s managed list.

Which assistants actually reach a small business site

August 2026 · share of sites reached

Each bar is the share of monitored sites that assistant’s crawler reached. Reach rather than volume, because a single busy site can account for most of the requests in a month, and a request count would then tell you the opposite of what is happening.

ChatGPT-User
95.3%
OAI-SearchBot
87.9%
PerplexityBot
63.2%
YouBot
44.1%
Claude-User
44%
Perplexity-User
37.1%
GoogleOther
25.5%
Claude-SearchBot
23.8%
Why Google looks absent

Google is low here because its assistants mostly do not fetch. AI Overviews and AI Mode answer from the Search index Googlebot has already taken, so there is usually no second request to log: Googlebot reaches almost every site in this fleet and makes roughly a thousand requests for every one made by a Google agent our classifier labels as AI. Google-Extended, which people look for in this table, is a consent switch over training and grounding rather than a crawler. So for Google the lever is what you allow Googlebot and Google-Extended to do, not whether a Google AI crawler turns up.

Optimization leaderboardAugust 2026

Which single change moves the assistants?

We make several changes to a page at once, so the useful question is which of them is doing the work. This section ranks them. On each site we take pages of the same type, compare the ones that got a given change against the ones that did not, and count the sites where the change comes out ahead on AI referral visits. We re-run it every month, because what earns a citation keeps moving.

The changes themselves are ordinary. A five-question question-and-answer section, written from the page’s own content and marked up so an assistant can lift it, is ahead on 79% of sites. A title tag rewritten to name what the page actually offers, in the words someone would ask in, is ahead on 77%. Rewriting dense paragraphs so each one leads with its answer instead of building up to it, 69%. A single main heading that names the subject rather than greeting the visitor, 65%.

And one change we do not make, because it does nothing: llms.txt, the file this field has recommended for two years. AI crawlers fetched it 31 times across 82 million requests in six months.

How this is measured

Sites where the change comes out aheadAI referrals, 90d
ChangeSites won
FAQ block130 of 164 sites
79%
Meta title119 of 154 sites
77%
Heading and body structure126 of 184 sites
69%
H1 tag103 of 159 sites
65%
llms.txtno effect found
none

Running experiments

3 in flight

Everyone in this field reports correlations, because observing model answers is all their data allows. A plugin install is a dated change to a real website, which makes before and after possible.

Running · Q4

Does structured business data bring the bots back sooner?

Diff-in-differences · matched control cohort · 90-day pre/post

Running · Q4

Which pages do AI crawlers actually re-read, and how often?

Per-URL crawl frequency · top vs bottom decile · 14 attributes

Running · Q1

How volatile is AI crawling, week to week?

Same balanced panel · dispersion, not level

Methods

Every figure here is a share, a median or an index.

01

Cohorts are fixed and windows matched period over period, so a number never rises just because the panel did.

02

Medians carry the trend wherever a single heavy crawler or busy site would dominate a mean, which on this data is most of the time.

03

Partial periods are excluded or labelled. Never silently included.

04

Nothing identifies a site, a page or a customer.

05

“AI crawlers” means search and on-demand fetchers. Training crawlers are excluded: one alone makes more requests than every AI search bot combined, and tells you nothing about visibility.

16,662 scans

Checking AI search readiness

The readiness score is a 0-100 check of whether a page states its facts in a form an assistant can lift: structured business data, a single clear heading, answerable copy, machine-readable contact and service detail, and nothing in robots.txt or the server config turning the crawlers away. It measures whether you can be quoted, not whether you rank.

Under 50 14.9%50–59 25.0%60–69 42.8%70–79 14.7%80+ 2.7%

Last 30 days · 1,494 scans · median 62. A near-zero score, which is 5.2% of the window, can mean we could not read the site rather than that it reads badly, so the bottom band is partly measurement.

LovedByAI/Lab

Open research on AI search, measured from our own server logs and refreshed as new months close. Free to read, quote and cite.

© 2026 LovedByAI, Inc.