# AI crawler access test: can answer engines read these vendors?

> 4 of 4 tool sites reachable from our build machine on 2026-09-20. 0 of 4 restrict any checked AI crawler in robots.txt. 4 of 4 ship an /llms.txt (fireflies, otter, granola, tldv). Median homepage TTFB: fireflies fastest at 36ms, granola slowest at 414ms.

Human page: https://testedactually.com/ai-meeting-notes/benchmarks/ai-meeting-notes-ai-crawler-access/
Mirror: https://testedactually.com/ai-meeting-notes/benchmarks/ai-meeting-notes-ai-crawler-access.md
Updated: 2026-09-20

Methodology: https://testedactually.com/ai-meeting-notes/benchmarks/ai-meeting-notes-ai-crawler-access/ — script: `scripts/crawler-test.mjs`, n=4 tools × 3 runs, run 2026-09-20.

**AI crawler access test: can answer engines read these vendors? — n=4 tools, 3 runs per tool**

| Tool | TTFB ms (median) | HTML KB | /llms.txt | AI crawler policy | Blocked bots |
| --- | --- | --- | --- | --- | --- |
| [fireflies](/ai-meeting-notes/fireflies/) | 36 | 917 | yes | open | none |
| [otter](/ai-meeting-notes/otter/) | 239 | 208 | yes | open | none |
| [granola](/ai-meeting-notes/granola/) | 414 | 344.3 | yes | open | none |
| [tldv](/ai-meeting-notes/tldv/) | 59 | 372.5 | yes | open | none |

### Methodology

- Fetch https://<domain>/robots.txt with a plain HTTP GET and a desktop user agent.
- Parse User-agent groups line by line; a bot is blocked only if a group naming it (or the * wildcard) contains a 'Disallow: /' rule.
- Bots checked: GPTBot, ClaudeBot, CCBot, PerplexityBot, Google-Extended, Applebot-Extended, Bytespider.
- Policy: blocked = GPTBot, ClaudeBot, and CCBot all disallowed from /; partial = at least one checked bot disallowed; open = no restrictions; unknown = robots.txt unreachable.
- Fetch https://<domain>/llms.txt; llmsTxt = HTTP 200.
- Fetch the homepage 3 times; TTFB is time to response headers; report the median. HTML size is the decoded body length of the last successful fetch.

### Limitations

- Single machine, single network location — TTFB is directional, not global CDN truth.
- n=3 runs per tool; we report the median, not a distribution.
- Simplified robots.txt parsing: bot names matched case-insensitively, only Disallow rules read, path wildcards beyond 'Disallow: /' not modeled.
- Homepage only; vendor pricing pages may have different crawler policies.
- Snapshot of 2026-09-20; vendors change crawler policies without notice.
