Tim Hey

What agents read

Every machine-readable file this site serves, in one place. Click any of them. This is exactly what an agent sees when it visits.

curl timhey.co/.well-known/

GET /llms.txt
site map for models
GET /.well-known/ai-catalog.json
ARD capability catalog
GET /feed.xml
writing feed (RSS)
GET /robots.txt
crawl + AI-bot rules
GET /sitemap.xml
every URL, dated
GET /posts/the-ask-the-model-eval.md
markdown mirror, one per post

Every post ships two more an agent can use: the Markdown mirror above, and JSON-LD in the page head. View source on any post to read the typed facts an agent lifts instead of guessing.

New to this? Start with the files an agent checks before it reads your page.

Who’s crawling this, live

This site logs its own agent readers. Crawlers and agents don’t run JavaScript, so analytics never sees them — these counts come from the middleware, before the cache. It’s the argument these posts make, measuring itself.

6,050 agent visits from 21 distinct agents.

Not counted: 1,984 vulnerability scans — requests for .env files, PHP shells, and SSH keys this site has never served. Every public host gets them. They’re filtered out on purpose, because a number that counts scanners as readers isn’t measuring anything.

Also not counted: 2 requests from plain HTTP clientscurl, python-requests. Some of that is me testing this site with curl. It was landing in the agent column until I looked.

By agent

  • ByteDance2,430
  • Anthropic ClaudeBot1,188
  • unknown924
  • Googlebot257
  • OpenAI SearchBot254
  • Bingbot240
  • Meta171
  • Amazon145
  • OpenAI GPTBot138
  • OpenAI ChatGPT-User79
  • Common Crawl75
  • Applebot62
  • Perplexity34
  • Anthropic Claude-User18
  • Diffbot9
  • Applebot-Extended7
  • Cohere7
  • Perplexity-User5
  • Anthropic4
  • DuckDuckGo2
  • Anthropic Claude-Web1

Most requested paths

  • /robots.txt1,219
  • /661
  • /sitemap.xml634
  • /agents480
  • /resume472
  • /feed.xml139
  • /posts/the-registry-is-the-search-index134
  • /posts/the-agent-ready-sdk126

Volume

Aug 3last 14 daysAug 16

Recent hits

  • Anthropic ClaudeBot/agents08-16 19:23
  • Anthropic ClaudeBot/08-16 19:13
  • Anthropic ClaudeBot/sitemap.xml08-16 18:39
  • Anthropic ClaudeBot/robots.txt08-16 18:39
  • Googlebot/posts/the-agent-ready-sdk08-16 18:32
  • Googlebot/robots.txt08-16 18:28
  • Amazon/llms.txt08-16 17:47
  • Amazon/posts/the-trust-gap08-16 17:47
  • Amazon/posts/the-trust-gap08-16 17:47
  • Anthropic ClaudeBot/posts/a-standard-is-not-adoption08-16 17:25
  • OpenAI ChatGPT-User/08-16 16:43
  • Anthropic ClaudeBot/sitemap.xml08-16 16:29

Still unidentified

User-agents that asked for a machine-readable file but don’t match anything I recognize. This is the honest version of “unknown” — the list I work from when I add a new agent to the classifier.

  • Mozilla/5.0 (iPhone; CPU iPhone OS 13_2_3 like Mac OS X) AppleWebKit/605.1.15 (KHTML, like Gecko) Version/13.0.3 Mobile/83
  • Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/124.0.0.0 Safari/537.3660
  • Mozilla/5.0 (Linux; Android 7.0;) AppleWebKit/537.36 (HTML, like Gecko) Mobile Safari/537.36 (compatible; PetalBot;+http48
  • Mozilla/5.0 (Windows NT 10.0; Win64; x64) AppleWebKit/537.36 (KHTML, like Gecko) Chrome/107.0.0.0 Safari/537.3629
  • Mozilla/5.0 (compatible; Barkrowler/0.9; +https://babbar.tech/crawler)17
  • Mozilla/5.0 (compatible; SemrushBot/7~bl; +http://www.semrush.com/bot.html)13
  • Chrome Privacy Preserving Prefetch Proxy12
  • FindFiles.net/1.0 (compatible; +https://findfiles.net/bot)11

Raw JSON: /api/agent-traffic