ShowUp Labs / Guides

ShowUpLabsBot

You have probably arrived here from a line in your access logs. This page tells you exactly what our crawler does, and how to stop it in one line if you would rather it did not.

Mozilla/5.0 (compatible; ShowUpLabsBot/1.0; +https://showuplabs.com/bot)
It fetches a small number of public pages to check whether a business is mentioned on them. It obeys robots.txt. It waits at least four seconds between requests to the same host, and honours Crawl-delay when you set one. It does not train models on your content, it does not fetch anything behind a login, and it does not fetch your whole site. To block it entirely, add two lines to robots.txt and we will stop within a day.

Stop it right now

Add this to /robots.txt. It takes effect on our next run, within about a day, and we cache your robots file for the length of a single run rather than indefinitely.

User-agent: ShowUpLabsBot
Disallow: /

You do not have to email us or fill in a form. We check before we fetch, so a disallowed page is one we never request, rather than one we request and discard. If you would rather tell a human, the contact form works too, but the robots line is faster and does not depend on us reading anything.

What it actually does

Two jobs, both small.

What it does not do: it does not crawl your site (typically one URL per domain per week), it does not follow every link it finds, it does not fetch anything requiring a login, it does not submit forms, it ignores anything behind a paywall rather than working around it, and your content is never used to train a model. We do not train models.

Rate

Why we are careful about this

We sell crawler-access auditing. Our own product tells businesses whether AI crawlers can reach them and which ones they are blocking. Running a badly-behaved crawler while selling that would be indefensible, so the standard we hold ourselves to here is the one we would apply to anyone else.

There is a wider reason too. Bot and crawler traffic now exceeds human traffic on the web, and sites with something to lose have reasonably started shutting the door: roughly 60% of high-credibility news sites now block AI crawlers, and long-standing testing organisations disallow the major AI agents outright. That is a rational response to being scraped, and the crawlers that caused it are the ones that took without asking. We would rather be the kind that asks.

If we have got something wrong

If ShowUpLabsBot is hitting you more than described, ignoring your robots.txt, or appearing where you did not expect it, tell us and we will fix it and say what happened. We would rather be corrected than be confidently wrong, which is the same offer on every report we publish.

Questions, answered

How do I block ShowUpLabsBot?

Add two lines to your robots.txt: User-agent: ShowUpLabsBot then Disallow: / on the next line. It takes effect on our next run, within about a day. We check robots.txt before fetching, so a disallowed page is never requested rather than requested and discarded. No email or form is required.

Does ShowUpLabsBot use my content to train AI models?

No. We do not train models. The crawler fetches a page, checks whether a particular business name appears in the text, and keeps that yes or no answer along with the page's author and contact details where it is looking for those. Your page content is not stored and is not used for training.

How often does ShowUpLabsBot visit?

Typically one page per domain, about once a week, with a minimum of four seconds between requests to the same host and longer if your robots.txt sets a Crawl-delay. If you are seeing substantially more than that from us, it is a bug and we would like to hear about it.

Does ShowUpLabsBot respect robots.txt?

Yes, and it checks before fetching rather than after. It reads the group for ShowUpLabsBot if you have one and falls back to the wildcard group otherwise, applying longest-match-wins between Allow and Disallow as the standard specifies. If robots.txt returns a server error we stay out, because an ambiguous answer should not be treated as permission.

Why is a crawler from an SEO company fetching my page?

We measure which third-party pages influence what AI engines say about a business category, then check whether a given company is named on them. Roughly 85% of what an AI engine cites is somebody else's page, so working out which pages those are means reading them. It is one fetch, it is a public page, and it obeys your robots.txt.

Curious what AI says about your site?

The same engine that powers this crawler will scan your own domain free and tell you which AI crawlers you are currently blocking, often without meaning to. About a minute, no signup.

Get your free score

Free score in under a minute. Live AI answers, not estimates. No signup, no card.