AI crawler access on .fm domains
7 of the 31 domains measured in this segment refuse a request from an identified AI user agent at the edge, 23% of the segment, against 3% whose robots.txt disallows an answer engine at the site root. The block that costs these sites citations is not the one written in the file. That answer-engine rate is 3.3 points below the corpus rate of 7%, which is itself measured across 27,313 domains.
Mean AI access score across those 31 domains is 70.6 of 100, 3 of them publish an llms.txt (10%), and 13 publish structured data an answer engine can parse (42%). At 31 domains this segment is small enough that one site changing its robots.txt moves every rate on this page.
Block rate by crawler, inside this segment robots.txt at the site root
| Crawler | Operator | Content used for | Block rate | Blocking | Measured |
|---|---|---|---|---|---|
| Bytespider | ByteDance | training | 19.4% | 6 | 31 |
| GPTBot | OpenAI | training | 16.1% | 5 | 31 |
| Meta-ExternalAgent | Meta | training | 12.9% | 4 | 31 |
| Google-Extended | training | 12.9% | 4 | 31 | |
| ClaudeBot | Anthropic | training | 12.9% | 4 | 31 |
| CCBot | Common Crawl Foundation | archive | 12.9% | 4 | 31 |
| Applebot-Extended | Apple | training | 12.9% | 4 | 31 |
| Amazonbot | Amazon | training | 12.9% | 4 | 31 |
| SemrushBot-OCOB | Semrush | training | 3.2% | 1 | 31 |
| Meta-WebIndexer | Meta | answer index | 3.2% | 1 | 31 |
| YouBot | You.com | answer index | 0.0% | 0 | 31 |
| Timpibot | Timpi | training | 0.0% | 0 | 31 |
| TikTokSpider | ByteDance | live retrieval | 0.0% | 0 | 31 |
| Scrapy | Zyte (open-source framework) | training | 0.0% | 0 | 31 |
| PerplexityBot | Perplexity | answer index | 0.0% | 0 | 31 |
| Perplexity-User | Perplexity | live retrieval | 0.0% | 0 | 31 |
| PanguBot | Huawei | training | 0.0% | 0 | 31 |
| omgilibot | Webz.io | training | 0.0% | 0 | 31 |
| omgili | Webz.io | training | 0.0% | 0 | 31 |
| OAI-SearchBot | OpenAI | answer index | 0.0% | 0 | 31 |
| OAI-AdsBot | OpenAI | live retrieval | 0.0% | 0 | 31 |
| msnbot | Microsoft | answer index | 0.0% | 0 | 31 |
| MistralAI-User | Mistral AI | live retrieval | 0.0% | 0 | 31 |
| MistralAI-Training | Mistral AI | training | 0.0% | 0 | 31 |
| MistralAI-Index | Mistral AI | answer index | 0.0% | 0 | 31 |
| Meta-ExternalFetcher | Meta | live retrieval | 0.0% | 0 | 31 |
| Meta-ExternalAds | Meta | training | 0.0% | 0 | 31 |
| Kangaroo Bot | Kangaroo LLM | training | 0.0% | 0 | 31 |
| ImagesiftBot | ImageSift (Hive) | training | 0.0% | 0 | 31 |
| GoogleOther | training | 0.0% | 0 | 31 | |
| Googlebot-News | answer index | 0.0% | 0 | 31 | |
| Googlebot | answer index | 0.0% | 0 | 31 | |
| Google-CloudVertexBot | live retrieval | 0.0% | 0 | 31 | |
| facebookexternalhit | Meta | live retrieval | 0.0% | 0 | 31 |
| FacebookBot | Meta | training | 0.0% | 0 | 31 |
| DuckAssistBot | DuckDuckGo | live retrieval | 0.0% | 0 | 31 |
| Diffbot | Diffbot | training | 0.0% | 0 | 31 |
| cohere-training-data-crawler | Cohere | training | 0.0% | 0 | 31 |
| cohere-ai | Cohere | live retrieval | 0.0% | 0 | 31 |
| Claude-User | Anthropic | live retrieval | 0.0% | 0 | 31 |
| Claude-SearchBot | Anthropic | answer index | 0.0% | 0 | 31 |
| ChatGPT-User | OpenAI | live retrieval | 0.0% | 0 | 31 |
| Bingbot | Microsoft | answer index | 0.0% | 0 | 31 |
| Applebot | Apple | answer index | 0.0% | 0 | 31 |
| anthropic-ai | Anthropic | training | 0.0% | 0 | 31 |
| Amzn-User | Amazon | live retrieval | 0.0% | 0 | 31 |
| Amzn-SearchBot | Amazon | answer index | 0.0% | 0 | 31 |
| AI2Bot | Allen Institute for AI | training | 0.0% | 0 | 31 |
Block rate is the share of this segment's domains holding a policy record for that agent whose robots.txt disallows it at the site root, shown with its numerator and denominator. Denominators differ between agents because a record only exists once a domain has been measured for it. A domain serving no robots.txt counts as allowing everything, which is what the standard specifies.
Most accessible
| Domain | Score | Engines blocked | llms.txt | |
|---|---|---|---|---|
| 1 | fireside.fm | 96 | 0 | yes |
| 2 | captivate.fm | 92 | 0 | no |
| 3 | transistor.fm | 91 | 0 | no |
| 4 | tupi.fm | 90 | 0 | no |
| 5 | zeno.fm | 90 | 0 | yes |
| 6 | riverside.fm | 84 | 0 | no |
| 7 | marus.fm | 84 | 0 | no |
| 8 | radiocut.fm | 83 | 0 | no |
| 9 | shorturl.fm | 76 | 0 | no |
| 10 | last.fm | 76 | 0 | no |
| 11 | castbox.fm | 75 | 0 | no |
| 12 | gayporno.fm | 73 | 0 | no |
| 13 | mel.fm | 71 | 0 | no |
| 14 | stats.fm | 71 | 0 | no |
| 15 | setlist.fm | 70 | 0 | no |
Highest AI access scores among the 31 domains measured in this segment.
Least accessible
| Domain | Score | Engines blocked | llms.txt | |
|---|---|---|---|---|
| 1 | podbay.fm | 45 | 0 | no |
| 2 | omny.fm | 53 | 0 | no |
| 3 | nic.fm | 56 | 0 | no |
| 4 | earn.fm | 57 | 0 | no |
| 5 | files.fm | 59 | 0 | yes |
| 6 | overcast.fm | 59 | 0 | no |
| 7 | radio.fm | 59 | 0 | no |
| 8 | anchor.fm | 62 | 0 | no |
| 9 | mobcup.fm | 62 | 0 | no |
| 10 | rci.fm | 62 | 1 | no |
| 11 | reaper.fm | 63 | 0 | no |
| 12 | znaki.fm | 64 | 0 | no |
| 13 | laut.fm | 64 | 0 | no |
| 14 | rmf.fm | 67 | 0 | no |
| 15 | dice.fm | 67 | 0 | no |
Lowest AI access scores among the same 31 domains.
Compare with the rest of the grouping top-level domain segments
The largest 12 other segments in this grouping, ordered by size. The share beside each is that segment's own answer-engine block rate over its own count. Full table on the segments index.
What this segment does not tell you read this before citing it
A top-level domain is a weak proxy for jurisdiction. Most of them are open to any registrant anywhere, so a domain under a country-code registry may be operated from outside that country by a company no local law reaches. Read these rows as differences between registration communities and the publishing cultures inside them, not as differences between legal regimes.
Measure your own domain
The audit that produced every number on this page runs on any domain in about six seconds: robots.txt resolved for 48 agents, then live requests sent as four of them to see whether the edge agrees with the file.