Identification
Its full user-agent string, exactly as it arrives at the gate:
There is no operator to verify against. This name is what software calls itself when nobody gave it one. Read fossick.bot's documentation.
How to block FossickBot
Add these two lines to the robots.txt file at the root of your site. Well-behaved crawlers read it before they crawl, so the change applies from FossickBot's next visit. Nothing else on your site needs to change.
# Block FossickBot from the whole site User-agent: FossickBot Disallow: /
Or let it visit but keep it away from part of the site:
# Let it in, but keep it out of one room User-agent: FossickBot Allow: / Disallow: /members/
Observed behaviour
Most active around 19:00. After dark, like a burglar.
Requested 0 disallowed pages out of 5 requests. Read robots.txt 2 times.
Has never followed the hidden link to /trap/. Either well trained or very lucky.
Where it comes from
Scripts and scanners run from wherever their owners rent a server. These are the networks behind the visits on file:
Networks and countries come from the visitor's IP address, looked up in a local copy of the DB-IP database. The addresses themselves are never stored.
Keeper's field notes
Questions site owners ask
Does FossickBot respect robots.txt?
Yes. In 5 requests observed here it has read robots.txt and never fetched a disallowed page.
Will blocking FossickBot hurt my search rankings?
No. Nothing respectable will miss it.
How often does FossickBot visit?
Here, about 1 requests a day over the last week. Visits to your site depend on its size, how often it changes, and how many links point to it.