Does yelp.com allow CCBot? robots.txt as of 2026-09-10
On 2026-09-10 the robots.txt of yelp.com resolved 19 AI crawler tokens: 2 allowed for / and 17 disallowed for /.
| Domain | yelp.com |
|---|---|
| AI crawler tokens resolved | 19 |
| Allowed for / | 2 |
| Disallowed for / | 17 |
| robots.txt | robots.txt present, 11825 bytes |
| Checks that ran | robots |
| robots.txt read (tier A) | 2026-09-10 |
| Full check (tier B) | 2026-09-10 |
| Last verified |
On this domain the robots.txt only check (tier A) last ran on 2026-09-10 and the full check (tier B) last ran on 2026-09-10. The full check adds the home page, the labelled edge probes, llms.txt, the headers and the structured data; the robots.txt check is the one that runs weekly for every seeded domain.
One question per crawler
- GPTBot
- ClaudeBot
- PerplexityBot
- OAI-SearchBot
- CCBot
- Bytespider
- Googlebot
- ChatGPT-User
- Google-Extended
Does yelp.com allow GPTBot? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 333 in the User-agent: GPTBot group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow GPTBot
User-agent: GPTBot Allow: /
User-agent: GPTBot Disallow: /
The GPTBot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow ClaudeBot? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 333 in the User-agent: ClaudeBot group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow ClaudeBot
User-agent: ClaudeBot Allow: /
User-agent: ClaudeBot Disallow: /
The ClaudeBot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow PerplexityBot? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 322 in the User-agent: PerplexityBot group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow PerplexityBot
User-agent: PerplexityBot Allow: /
User-agent: PerplexityBot Disallow: /
The PerplexityBot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow OAI-SearchBot? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 312 in the User-agent: OAI-SearchBot group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow OAI-SearchBot
User-agent: OAI-SearchBot Allow: /
User-agent: OAI-SearchBot Disallow: /
The OAI-SearchBot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow CCBot? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 333 in the User-agent: CCBot group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow CCBot
User-agent: CCBot Allow: /
User-agent: CCBot Disallow: /
The CCBot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow Bytespider? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow Bytespider
User-agent: Bytespider Allow: /
User-agent: Bytespider Disallow: /
The Bytespider guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow Googlebot? robots.txt as of 2026-09-10
allowed for /
robots.txt has a User-agent: Googlebot group with no rule matching / , so / is allowed. 3 records name this user agent and RFC 9309 merges them
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow Googlebot
User-agent: Googlebot Allow: /
User-agent: Googlebot Disallow: /
The Googlebot guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow ChatGPT-User? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 322 in the User-agent: ChatGPT-User group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow ChatGPT-User
User-agent: ChatGPT-User Allow: /
User-agent: ChatGPT-User Disallow: /
The ChatGPT-User guide: identity, purpose, documented verification and the vendor sources.
Does yelp.com allow Google-Extended? robots.txt as of 2026-09-10
disallowed for /
robots.txt line 333 in the User-agent: Google-Extended group: "Disallow: /" (longest match wins)
Observed 2026-09-10, first observed 2026-09-10, last change recorded 2026-09-10.
Allow or disallow Google-Extended
User-agent: Google-Extended Allow: /
User-agent: Google-Extended Disallow: /
The Google-Extended guide: identity, purpose, documented verification and the vendor sources.
Every crawler token in the registry
| Crawler | Vendor | Observation | Evidence | Observed |
|---|---|---|---|---|
| GPTBot | OpenAI | disallowed for / | robots.txt line 333 in the User-agent: GPTBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| ClaudeBot | Anthropic | disallowed for / | robots.txt line 333 in the User-agent: ClaudeBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| PerplexityBot | Perplexity | disallowed for / | robots.txt line 322 in the User-agent: PerplexityBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| OAI-SearchBot | OpenAI | disallowed for / | robots.txt line 312 in the User-agent: OAI-SearchBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| CCBot | Common Crawl Foundation | disallowed for / | robots.txt line 333 in the User-agent: CCBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Bytespider | ByteDance | disallowed for / | robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Googlebot | allowed for / | robots.txt has a User-agent: Googlebot group with no rule matching / , so / is allowed. 3 records name this user agent and RFC 9309 merges them | 2026-09-10 | |
| ChatGPT-User | OpenAI | disallowed for / | robots.txt line 322 in the User-agent: ChatGPT-User group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Claude-User | Anthropic | disallowed for / | robots.txt line 322 in the User-agent: Claude-User group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Claude-SearchBot | Anthropic | disallowed for / | robots.txt line 322 in the User-agent: Claude-SearchBot group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Perplexity-User | Perplexity | disallowed for / | robots.txt line 322 in the User-agent: Perplexity-User group: "Disallow: /" (longest match wins) | 2026-09-10 |
| anthropic-ai | Anthropic | disallowed for / | robots.txt line 333 in the User-agent: anthropic-ai group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Bingbot | Microsoft | allowed for / | robots.txt has a User-agent: Bingbot group with no rule matching / , so / is allowed. 2 records name this user agent and RFC 9309 merges them | 2026-09-10 |
| Applebot | Apple | disallowed for / | robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Amazonbot | Amazon | disallowed for / | robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins) | 2026-09-10 |
| meta-externalagent | Meta | disallowed for / | robots.txt line 333 in the User-agent: meta-externalagent group: "Disallow: /" (longest match wins) | 2026-09-10 |
| DuckAssistBot | DuckDuckGo | disallowed for / | robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Applebot-Extended | Apple | disallowed for / | robots.txt line 336 in the User-agent: * group: "Disallow: /" (longest match wins) | 2026-09-10 |
| Google-Extended | disallowed for / | robots.txt line 333 in the User-agent: Google-Extended group: "Disallow: /" (longest match wins) | 2026-09-10 |
What else the robots.txt says
| agents-summary | AI crawler rules: 2 allowed, 17 disallowed 19 crawler tokens resolved against https://yelp.com/robots.txt (HTTP 200, 11825 bytes) at 2026-09-10T23:46:59Z |
|---|---|
| file | robots.txt present, 11825 bytes GET https://yelp.com/robots.txt at 2026-09-10T23:46:59Z returned HTTP 200, 11825 bytes, content type text/plain, 9 user agent group(s) |
| our-checker | robots.txt disallows our checker robots.txt line 336: "Disallow: /" applies to BikooshBot/1.0 (+https://bikoosh.com/bot); nothing further was fetched |
History
Every robots.txt version we have observed for yelp.com.
Check your own site
The same checks, run on your domain now: what robots.txt tells each AI crawler, whether the edge answers them, and whether the text is readable without JavaScript. Free, no signup.
Limits: 8 checks and 5 different domains per address per hour; a repeat of the same domain inside 7 days is answered from what we already hold.
Limitation: this test appends a labelled marker to our own honest User Agent, it never impersonates another agent. Rules keyed to an exact User Agent string, or to a vendor's IP ranges, are invisible to this method and are neither confirmed nor ruled out.
AEO Watch is an independent, factual monitor. It is not affiliated with, endorsed by or speaking for any crawler vendor. Every statement is an observation with the date it was made and the raw evidence behind it: a robots.txt line, an HTTP status code, a header. There are no scores, no grades and no verdicts here, and nothing on this page is advice. A site opts out at any time and the opt out is honoured automatically and permanently.