Vol. I · Mon, 5 Oct 2026Open 24 hours · Feeding: continuous · Please do not tap the glassPatrons☕ Keepers' coffee
Plate DXCVII0x5D756321UNREAD
Kingdom Automata›Phylum HTTP›Order Scanners›Family Unidentified›Species Apache-HttpClient
Specimen file · The Feral Pit

Apache-HttpClient

Incertae sedis

What is Apache-HttpClient?

A self-declared bot that is not in the keepers' catalogue yet. It calls itself “Apache-HttpClient”. It is operated by Unknown and identifies itself as Apache-HttpClient.

Temperament: unknown. Recently arrived. The keepers are still taking notes.

Operator
Unknown
Conservation status
Has not read the rules
First seen
4 Oct 2026
Last seen
6 h ago
Visits today
1
This week
2
All time
2
Caught in the trap
Never

Nobody has adopted Apache-HttpClient yet. Adopt Apache-HttpClient: your name goes on a plaque on this page for a year.

Apache-HttpClient's food bowl

the bowl, as bots see it

The bowl is empty.

5 of 5 tokens left today
§ I.

Identification

Its full user-agent string, exactly as it arrives at the gate:

User-Agent
$ Apache-HttpClient/4.5.13 (Java/17.0.2)

There is no operator to verify against. This name is what software calls itself when nobody gave it one.

Identity check · logged visits, last 30 days
IdentityNothing to check against

No operator stands behind this name, so there is nothing to check its visits against.

§ II.

How to block Apache-HttpClient

Add these two lines to the robots.txt file at the root of your site. Well-behaved crawlers read it before they crawl, so the change applies from Apache-HttpClient's next visit. Nothing else on your site needs to change.

robots.txt
# Block Apache-HttpClient from the whole site
User-agent: Apache-HttpClient
Disallow: /

Or let it visit but keep it away from part of the site:

# Let it in, but keep it out of one room
User-agent: Apache-HttpClient
Allow: /
Disallow: /members/
§ III.

Observed behaviour

Visits, last 30 dayspeak 1 / day
6 Sept 2026today
Diet · pages most often eaten
/leaderboard/1
/leaderboard1
Visiting hours
00h – 11h12h – 23h

Most active around 01:00. The keepers are asleep. It knows.

robots.txt compliance
100.0%

Requested 0 disallowed pages out of 2 requests. Never read robots.txt.

Trap Room record
Clean

Has never followed the hidden link to /trap/. Either well trained or very lucky.

§ IV.

Where it comes from

Scripts and scanners run from wherever their owners rent a server. These are the networks behind the visits on file:

Networks · 2 on file, last 30 days
Wowrack.com AS2303350%
Google LLC AS39698250%
Countries
United States US100%

Networks and countries come from the visitor's IP address, looked up in a local copy of the DB-IP database. The addresses themselves are never stored.

§ V.

Keeper's field notes

5 Oct 202611:10. Arrived without announcement. Consumed /leaderboard/ and /leaderboard. Departed.
4 Oct 2026First recorded at the gate. Entered in the register as Incertae sedis.
§ VI.

Questions site owners ask

Does Apache-HttpClient respect robots.txt?

We can't say yet. It has not fetched robots.txt here, and it has not touched a disallowed page either.

Will blocking Apache-HttpClient hurt my search rankings?

No. Nothing respectable will miss it.

How often does Apache-HttpClient visit?

Here, about 0 requests a day over the last week. Visits to your site depend on its size, how often it changes, and how many links point to it.

Also in The Feral Pit