Crawler policy

NetaRecordBot

This project fetches public records from government portals. If you operate one of them and our crawler is causing you trouble, here is everything you need to identify and stop it.

How to identify it

Our user agent is exactly:

NetaRecordBot/1.0 (+https://netarecord.org/about/bot; data@netarecord.org)

What it does

How to stop it

Add this to your robots.txt and it will stop within one run:

User-agent: NetaRecordBot
Disallow: /

Or email data@netarecord.org and we will add your host to the block list directly, which takes effect on the next deploy rather than the next run.

What it never does

It does not attempt to log in, does not submit forms other than the public filter forms a portal exposes for browsing, does not follow links behind an authentication wall, and does not collect personal data beyond what appears in the public record it is reading. It does not fetch phone numbers or home addresses, and it stores constituency-level location at most.