Now tracking 178 AI crawlers & agents

See which AI bots are
reading your WordPress site

GPTBot, ClaudeBot, PerplexityBot and 175+ more — who visits, what they fetch, and whether your robots.txt covers them. No Cloudflare account, no guessing.

free forever core · no account needed · nothing leaves your site unless you opt in

178AI crawlers & agents in the fingerprint library
2open upstreams, merged and updated weekly
0data leaves your site — cloud is opt-in only

How it works

Up and running in one minute

Install like any plugin. CrawlTally starts counting immediately — no configuration, no account.

01

Install & activate

From your WordPress admin, Plugins → Add New. The fingerprint library ships inside the plugin — it works offline.

02

See your bots

“GPTBot made 1,204 requests this week (+38%)” — trends, first-seen dates and top crawled paths, in plain English.

03

Decide & act

Audit your robots.txt per AI crawler, then generate a new one in one click: block training bots, keep search bots citable.

Features

Everything a site owner needs

Built for WordPress long-tail sites — readable answers, not a security dashboard.

🔍

178-bot fingerprint library

Every known AI crawler and agent with vendor, purpose (training / search / agent) and official docs link. Updated weekly from two open upstreams — ships inside the plugin.

GPTBotClaudeBotPerplexityBotGoogle-ExtendedBytespiderCCBotGemini+171 more
🤖

robots.txt audit & generator

Per-bot verdicts and one-click generation with sane presets — “block training, allow search” recommended.

GPTBotblocked
ClaudeBotallowed
PerplexityBotnot mentioned
📈

Weekly email reports optional cloud

“This week: 6 AI bots, 3,200 requests. ClaudeBot visited for the first time.” Opt-in only, aggregated counters, no visitor IPs.

📡

Cache-proof census beacon

Page cache or CDN eating your numbers? The never-cached census endpoint keeps “which bots visit” measurable even behind full-page caches.

🗂️

Access-log import Pro

Behind aggressive caching? Import your Nginx / Apache access logs for cache-immune, byte-accurate history. Runs on your server via WP-CLI — the data never leaves it.

wp crawltally import-log
🧭

Honest by design

User-agent-level detection (and we say so), lower-bound counts behind caches (we label it), and we won’t sell you a llms.txt generator — independent measurements show it barely matters. You get the truth, not snake oil.

✓ we label what we measure ✓ cache-aware by default ✗ no llms.txt snake oil

Pricing

Simple, honest pricing

Start free. Upgrade only if you want it. No subscriptions unless you choose one.

FREE

$0
  • 7-day bot dashboard
  • robots.txt audit + generator
  • Census beacon
  • 3 index checks / week
  • 90-day history & export
  • Email reports
  • Log import
Download
★ Most popular

PRO

$79 one-time
  • Everything in Free
  • 90-day history
  • Cache-immune log import
  • Weekly email reports (self-hosted)
  • 20 index checks / week
  • 3 sites
  • 1 year fingerprint library updates

CLOUD

$9 / month
  • Everything in Pro
  • Managed weekly reports & first-seen alerts
  • Multi-site overview (3 sites)
  • Cancel anytime

Secure checkout by Creem (Merchant of Record) · VAT handled · License key arrives by email in minutes.

FAQ

Frequently asked

My site uses a page cache or CDN — are the numbers wrong?

Requests served from cache never reach PHP, so request counts are a lower bound on cached sites — CrawlTally detects your cache and says so in the dashboard. Which bots visit and first-seen dates stay reliable, and the census beacon plus Pro log import keep detection complete.

Why don't the numbers match my Cloudflare panel?

Cloudflare sees every request at the edge; a WordPress plugin sees what reaches your server. If Cloudflare caches your HTML, the difference is exactly your cache hit rate. Note that Cloudflare's default does not cache HTML — most sites see near-identical totals.

Do you generate llms.txt for me?

No. Independent measurements (500 million AI-bot visits, 408 llms.txt reads) show it has almost no effect on AI visibility. We check whether you have one and tell you the truth instead of selling snake oil.

What data leaves my site?

Nothing, by default. If you opt in to cloud reports, only aggregated counters do: bot ID, path section, hit counts. No visitor IPs, no full URLs, no content.

Can bots fake their user agent?

Yes — identification is user-agent based, and a spoofed UA counts as that bot. For most site owners this precision is exactly right; IP-range verification is on the roadmap.

Stop guessing who's reading your site

Free forever core, one-minute setup, honest numbers.

Download free from WordPress.org

works with any cache plugin · any CDN · any host