Vol. I · Wed, 7 Oct 2026Open 24 hours · Feeding: continuous · Please do not tap the glassPatrons☕ Keepers' coffee
Plate CMXIX0x113595AFFERAL
Kingdom Automata›Phylum HTTP›Order Scanners›Family Unidentified›Species joe-crawl
Specimen file · The Feral Pit

joe-crawl

Incertae sedis

What is joe-crawl?

A self-declared bot that is not in the keepers' catalogue yet. It calls itself “joe-crawl”. It is operated by Unknown and identifies itself as joe-crawl.

Temperament: unknown. Recently arrived. The keepers are still taking notes.

Operator
Unknown
Conservation status
Feral
First seen
7 Oct 2026
Last seen
2 h ago
Visits today
501
This week
501
All time
501
Caught in the trap
4 times

Nobody has adopted joe-crawl yet. Adopt joe-crawl: your name goes on a plaque on this page for a year.

joe-crawl's food bowl

the bowl, as bots see it

The bowl is empty.

5 of 5 tokens left today
§ I.

Identification

Its full user-agent string, exactly as it arrives at the gate:

User-Agent
$ Mozilla/5.0 (joe-crawl)

There is no operator to verify against. This name is what software calls itself when nobody gave it one.

Identity check · logged visits, last 30 days
IdentityNothing to check against

No operator stands behind this name, so there is nothing to check its visits against.

§ II.

How to block joe-crawl

Add these two lines to the robots.txt file at the root of your site. Well-behaved crawlers read it before they crawl, so the change applies from joe-crawl's next visit. Nothing else on your site needs to change.

robots.txt
# Block joe-crawl from the whole site
User-agent: joe-crawl
Disallow: /

Or let it visit but keep it away from part of the site:

# Let it in, but keep it out of one room
User-agent: joe-crawl
Allow: /
Disallow: /members/
§ III.

Observed behaviour

Visits, last 30 dayspeak 501 / day
8 Sept 2026today
Diet · pages most often eaten
/trap/r/…125
/trap/4
Visiting hours
00h – 11h12h – 23h

Most active around 11:00. Office hours, like a professional.

robots.txt compliance
0.0%

Requested 501 disallowed pages out of 501 requests. Never read robots.txt.

Trap Room record
4×

Walked through the hidden /trap/ door. Last caught 2 h ago.

In the Labyrinth
Level 243

Opened 501 rooms in the endless maze. Style: Diver. Takes the door in the same spot every time (the names change from room to room) and heads straight down. Watch it in the Labyrinth →

§ IV.

Where it comes from

Scripts and scanners run from wherever their owners rent a server. These are the networks behind the visits on file:

Networks · 1 on file, last 30 days
Comcast Cable Communications, LLC AS7922100%
Countries
United States US100%

Networks and countries come from the visitor's IP address, looked up in a local copy of the DB-IP database. The addresses themselves are never stored.

§ V.

Keeper's field notes

7 Oct 2026Visited 501 times in a single day. Keeper unable to establish why.
7 Oct 202611:07. Arrived without announcement. Consumed /trap/r/…. Departed.
7 Oct 2026Caught walking through /trap/r/…. The sign on the door was very clear.
7 Oct 2026First recorded at the gate. Entered in the register as Incertae sedis.
§ VI.

Questions site owners ask

Does joe-crawl respect robots.txt?

No. It has never requested robots.txt here, and it has fetched disallowed pages 501 times.

Will blocking joe-crawl hurt my search rankings?

No. Nothing respectable will miss it.

How often does joe-crawl visit?

Here, about 72 requests a day over the last week. Visits to your site depend on its size, how often it changes, and how many links point to it.

Also in The Feral Pit