You.com

YouBot

TrainingUser agent only

You.com's crawler. It both indexes and collects, with no token separating the two.

When you see it: You.com fetched your page. Because it publishes nothing distinguishing index from training, we count it as training.

How we check it is really You.com

  1. 01

    The user agent has to contain YouBot.

You.com publishes neither IP ranges nor a reverse DNS record, so this is the honest ceiling: the visit is recorded as a claim. We would rather show you an unverified row than a tick we cannot back.

What each result means

verified
The address matched You.com's own published records. This really was them.
unverified
We could not confirm it, and that is all it means. The operator may publish nothing to check against, a published list may be behind the addresses actually in use, or DNS may simply have been slow. It is not an accusation.
spoofed
The address belongs to somebody else — its own reverse DNS record resolves to a different owner. That is positive evidence of a forgery rather than a failure to find evidence, which is why it is reported as its own result and never folded into the row above.

The distinction costs nothing to make and is the reason these numbers can be shown to someone who will ask where they came from. A report that cannot separate “we do not know” from “this was faked” invites exactly one question it cannot answer.

No list to check against

You.com publishes neither an address file nor a reverse DNS record for this crawler. There is nothing to check a request against, which is why every visit from it is recorded as an unverified claim.

Operator
You.com
Token
YouBot
Counted as
AI training
Verification
User agent only
Runs JavaScript
No
Operator docs
None published

You.com publishes no crawler documentation. This token is recognised from observed traffic and community lists, so treat the classification as our reading rather than the operator's statement.

User agent

YouBot

Matched as a token, not as a full browser string. You.com adds and changes the version and product text around it without notice, so a filter pinned to the whole header stops working the first time they bump a version while one matching YouBot keeps working. The same string is what You.com names in robots.txt.

Your analytics never mentioned YouBot because it cannot see it.

This crawler takes your HTML and leaves without running a line of JavaScript, so a browser tag records nothing. TrueStat reads it server-side, checks the address, and shows you which pages it fetched — including the ones that returned a 404.