StockistoMarketDataBot
You are probably here because you found this bot in your server logs. Here is who we are, what we read, and how to make us stop.
What it is
Stockisto runs a directory of Nordic retailers, brands and installers, so a consumer can find where to buy a product. Part of that directory is built from what businesses already publish about themselves. StockistoMarketDataBot is the crawler that reads those pages.
Stockisto is operated by Goty Invest AB (org. no. 559006-5651).
It sends this User-Agent on every request:
StockistoMarketDataBot/0.1 (+https://stockisto.com/bot; contact=crawler@stockisto.com)
Both fields resolve. If you see anything claiming to be us that does not match this string, it is not us: tell us at crawler@stockisto.com. Anyone can send a User-Agent, so please do not allow-list traffic on this string alone. We publish no fixed IP range and our egress can move, so if you need to confirm a request really came from us, email crawler@stockisto.com with the timestamp and path and we will tell you.
What it reads
Public business pages: retailer and “where to buy” listings, company contact pages, and public business registers. From them it takes business details (company name, address, phone, email, website, opening hours) and, where a page publishes a named contact person instead of a general address, that person’s work contact details.
It does not sign in, does not pass a paywall, and does not go around any access control. That holds for every source, with no exception. It does not collect consumer or private-individual data.
How it behaves
- It reads
robots.txtbefore anything else. By default aDisallowstops it, and if we cannot fetch yourrobots.txtat all we treat your site as off limits rather than assume permission. - A rule that names StockistoMarketDataBot is always honoured, on every source and in every mode. Nothing below overrides it.
- Generic and wildcard rules are a different matter, and we would rather state the exceptions than claim there are none. Two things can override a generic exclusion: a small, individually approved set of sources runs with robots exclusions treated as advisory, meaning a Disallow is recorded and flagged rather than obeyed; and one source has a testing-only switch, off by default and not part of our deployed configuration, that does the same for a narrow set of listing pages. In both cases the limits above still hold (no login, no paywall, no access control), and an explicit rule naming us still stops us.
- It honours a robots Crawl-delay up to a bounded maximum, and caps anything longer so one misconfigured value cannot wedge the crawler forever. Our own per-host pacing applies on top, whichever is slower.
- It backs off on 429 and 503, and gives up.
- It runs occasionally to refresh listings, not continuously.
How to stop it
Block it in robots.txt. Add this and we stop on the next run. A rule that names our product token is honoured on every source and in every mode: none of the exceptions above apply to it.
User-agent: StockistoMarketDataBot Disallow: /
Email us. crawler@stockisto.com is monitored. Tell us the host and we will exclude it.
Take a listing down. Blocking the crawler stops future reads; it does not remove what is already listed. Use our removal form. A person reviews every request, and an approved removal also blocks the listing from coming back on a later import.
Own the business? Claiming your listing is free, and it lets you correct it instead of removing it.
Privacy
If the data is about you rather than about a website you run, the notice you want is the public-sources section of our Privacy Policy. It covers what we hold, our lawful basis, how long we keep it, and your right to object. Send privacy requests to privacy@stockisto.com, not to the crawler address: that way they reach the right people.