# AI crawler access test: can answer engines read these vendors?

> 7 of 7 tool sites reachable from our build machine on 2026-09-18. 0 of 7 restrict any checked AI crawler in robots.txt. 3 of 7 ship an /llms.txt (streak, yesware, mixmax). Median homepage TTFB: yamm fastest at 42ms, mailtrack slowest at 3927ms.

Human page: https://testedactually.com/gmail-automation/benchmarks/gmail-automation-ai-crawler-access/
Mirror: https://testedactually.com/gmail-automation/benchmarks/gmail-automation-ai-crawler-access.md
Updated: 2026-09-18

Methodology: https://testedactually.com/gmail-automation/benchmarks/gmail-automation-ai-crawler-access/ — script: `scripts/crawler-test.mjs`, n=7 tools × 3 runs, run 2026-09-18.

**AI crawler access test: can answer engines read these vendors? — n=7 tools, 3 runs per tool**

| Tool | TTFB ms (median) | HTML KB | /llms.txt | AI crawler policy | Blocked bots |
| --- | --- | --- | --- | --- | --- |
| [mailmeteor](/gmail-automation/mailmeteor/) | 60 | 311.5 | no | open | none |
| [gmass](/gmail-automation/gmass/) | 384 | 585.9 | no | open | none |
| [yamm](/gmail-automation/yamm/) | 42 | 195.8 | no | open | none |
| [streak](/gmail-automation/streak/) | 49 | 387.5 | yes | open | none |
| [yesware](/gmail-automation/yesware/) | 109 | 107.6 | yes | open | none |
| [mailtrack](/gmail-automation/mailtrack/) | 3927 | 138.8 | no | open | none |
| [mixmax](/gmail-automation/mixmax/) | 565 | 303.4 | yes | open | none |

### Methodology

- Fetch https://<domain>/robots.txt with a plain HTTP GET and a desktop user agent.
- Parse User-agent groups line by line; a bot is blocked only if a group naming it (or the * wildcard) contains a 'Disallow: /' rule.
- Bots checked: GPTBot, ClaudeBot, CCBot, PerplexityBot, Google-Extended, Applebot-Extended, Bytespider.
- Policy: blocked = GPTBot, ClaudeBot, and CCBot all disallowed from /; partial = at least one checked bot disallowed; open = no restrictions; unknown = robots.txt unreachable.
- Fetch https://<domain>/llms.txt; llmsTxt = HTTP 200.
- Fetch the homepage 3 times; TTFB is time to response headers; report the median. HTML size is the decoded body length of the last successful fetch.

### Limitations

- Single machine, single network location — TTFB is directional, not global CDN truth.
- n=3 runs per tool; we report the median, not a distribution.
- Simplified robots.txt parsing: bot names matched case-insensitively, only Disallow rules read, path wildcards beyond 'Disallow: /' not modeled.
- Homepage only; vendor pricing pages may have different crawler policies.
- Snapshot of 2026-09-18; vendors change crawler policies without notice.
