Skip to main content
MissedTakes
Menu

AI crawler traffic statistics: AI bots vs Googlebot on a sports and politics site

MissedTakes grades NFL and US election pundits’ predictions. Every request from a crawler that names itself in its user agent is logged as it reaches the site, and this page adds them up: AI crawlers next to Googlebot, by company, by purpose and by kind of page. It is one sports and politics data site’s logs, not the web.

Requests logged on MissedTakes from 2026-09-06 to 2026-10-05 (UTC), 30 days.

AI crawler requests
119,259

30 days to 2026-10-05

Googlebot requests
3,664

same days

AI vs Googlebot
33 times

as many requests

Top AI crawler
OAI-SearchBot

60,522 requests (OpenAI)

How many requests do AI crawlers make compared to Googlebot?

AI crawlers made 33 times as many requests as Googlebot: 119,259 against 3,664 in the 30 days to 2026-10-05. OAI-SearchBot alone made 17 times as many as Googlebot.

Which AI crawler crawls the most?

Download CSV

OAI-SearchBot, OpenAI's crawler, made the most requests of any AI crawler: 60,522 in the 30 days to 2026-10-05, 51% of all AI crawler requests. Next came ClaudeBot (42,542) and GPTBot (15,989).

By company, OpenAI's crawlers made 64% of AI crawler requests, Anthropic's 36% and Perplexity's under 1%.

Requests by crawler, 2026-09-06 to 2026-10-05
CrawlerCompanyTypeRequestsRefused (403)Share of bot requests
OAI-SearchBotOpenAIAI60,522—29%
Unrecognized botsUnknownOther56,521—27%
ClaudeBotAnthropicAI42,542—20%
YandexBotYandexSearch engine21,955—10%
GPTBotOpenAIAI15,989—8%
BaiduspiderBaiduSearch engine7,831—4%
GooglebotGoogleSearch engine3,664—2%
BingbotMicrosoftSearch engine1,846—1%
ChatGPT-UserOpenAIAI180—under 1%
ApplebotAppleSearch engine63—under 1%
PerplexityBotPerplexityAI20—under 1%
LivelapBotLivelapOther86under 1%
Claude-UserAnthropicAI6—under 1%
DuckDuckBotDuckDuckGoSearch engine3—under 1%

What share of crawler traffic is AI?

AI crawlers made 56% of the 211,150 requests we logged from bots in the 30 days to 2026-10-05; search engine crawlers made 17%, and 27% came from scrapers and bots that named no crawler we recognize.

Which crawlers does MissedTakes block?

Each ban followed a crawler that ignored robots.txt and fetched hard enough to strain the site; the reasons are on the record in our robots.txt. Search engines and AI crawlers that read robots.txt are all welcome.

LivelapBot (6) is banned on this site and answered with a 403 everywhere but robots.txt, so its requests in the 30 days to 2026-10-05 are attempts, not pages read.

Are AI crawlers collecting training data or answering users?

Each crawler is placed by what its operator says it is for. The operator's word is the only evidence of purpose a log has.

Crawlers their operators say collect training data made 49% of AI crawler requests; AI search indexers made 51%, and fetches made when a user asked an assistant under 1%.

AI crawler requests by stated purpose
Stated purposeCrawlersRequestsShare of AI requests
Collects training dataClaudeBot, GPTBot58,53149%
Builds an AI search indexOAI-SearchBot, PerplexityBot60,54251%
Fetches a page when a user asksChatGPT-User, Claude-User186under 1%

Which pages do AI crawlers read most?

Page families are read from the URL path, the same grouping our Search Console reports use. Only requests we served count here; a refused request read nothing.

Graded takes drew the most AI crawler requests, 29% of AI crawler requests, followed by NFL player pages (15%) and about and policy pages (14%). Googlebot's most-requested family was graded takes, at 39% of its requests.

Served AI crawler and Googlebot requests by page family
PagesAI crawlersShareGooglebotShare
Graded takes34,38229%1,43139%
NFL player pages17,37315%1915%
About and policy pages16,44514%14under 1%
Pundit records16,04713%36510%
NFL team pages7,3026%391%
NFL hub and leaderboard pages6,8336%732%
Publisher pages5,4905%1153%
Prediction-type pages4,1954%301%
Posts3,3093%742%
The home page1,9792%10under 1%
Robots.txt, sitemaps and llms.txt1,7591%1,09730%
Other URLs1,4861%17under 1%
NFL game pages1,3531%832%
Election race pages6541%702%
Other election pages6501%552%
The methodology2under 1%00%

Do AI crawlers read llms.txt?

No AI crawler requested /llms.txt or /llms-full.txt in the 30 days to 2026-10-05, against 710 AI crawler requests for robots.txt.

19,951 AI crawler requests were for the markdown version of a page (the same URL plus /md), 17% of all AI crawler requests.

Is AI crawling growing?

AI crawler requests rose 273% week over week: 63,137 in the week starting 2026-09-28, against 16,939 the week before. Across the last 8 complete weeks the weekly count ranged from 1,136 to 108,280.

The logs this page reads begin on 2026-08-03, 64 days of history; 8 complete weeks are compared here.

Requests by complete week (Monday to Sunday, UTC)
Week startingAI crawlersGooglebotAll bots
2026-08-10108,280671111,053
2026-08-171,1361,4143,801
2026-08-243,0781096,678
2026-08-312,7343837,956
2026-09-075,2879511,410
2026-09-1431,14012944,601
2026-09-2116,93912532,043
2026-09-2863,1372,798116,587

How this is counted

  • A request is counted when its user agent names a crawler we recognize, before the page is served and whatever we answer. The site’s built scripts and styles are not counted; pages, share-card images, feeds, sitemaps, robots.txt and llms.txt are. Ordinary browser visits are never recorded.
  • Banned crawlers are counted too: their requests are logged, then refused with a 403. jscrawler and LivelapBot are also denied by a firewall rule ahead of the site, so most of their attempts never reach this log at all.
  • User agents are self-reported, and anyone can claim to be a crawler. AI crawlers are counted as they name themselves. 3,655 of Googlebot's 3,664 requests (100%) came from Google's own 66.249.x.x crawler addresses; the rest used the Googlebot name from other addresses.
  • Google-Extended and Applebot-Extended are robots.txt switches, not crawlers, so they do not appear in logs. Google’s AI features read pages through Googlebot itself, so some Googlebot requests serve AI answers too; a log cannot tell which.
  • A markdown fetch here is a request for a page’s /md twin by its own address. An agent that asks for markdown at the page’s URL (with an Accept header) gets the twin too, but is logged as a request for the page.
  • MissedTakes is a small, specialized site: pundit records, graded takes, and NFL and election pages. A larger site, or one that blocks different crawlers, will see different numbers. This is one site, not the web.
  • “Unrecognized” is traffic that looks automated but names no crawler we know. Only totals by crawler and by kind of page are published here: no addresses, no URLs, no single visits. How the log is kept is in the site notice.
  • Whole UTC days, ending yesterday; re-read every few hours.

Cite these numbers

Free to reuse under CC BY 4.0 with credit and a link to this page. Please give the dates: the window moves every day. Figures for 2026-09-06 to 2026-10-05, read 2026-10-06.

MissedTakes, "AI crawler traffic statistics", missedtakes.co/ai-crawlers, accessed 2026-10-06. Requests from 2026-09-06 to 2026-10-05 (UTC).

The per-crawler table as CSV; the whole page as markdown.