NewzWyreBot

User-Agent: NewzWyreBot/1.0 (+https://newzwyre.com/crawler)

NewzWyreBot is the crawler of NewzWyre, a media-intelligence service operated by Chronic Internet LLC (P.O. Box 510072, Kealia, HI 96751, USA). It visits news and media websites to collect the professional contact details that outlets publish about their journalists and editors: names, titles, beats, work email addresses printed on the page, public social profiles, and links to recent bylines. It builds a database used for relevant media outreach. See the privacy notice for the legal basis and your rights.

What it fetches

Robots.txt, sitemaps (especially author sitemaps), outlet home, about, masthead, staff and author pages, and the occasional article to read a byline. It does not log in, submit forms, or fetch pages behind a paywall.

How it behaves

robots.txtAlways read first and cached for 24 hours. Rules for NewzWyreBot apply, otherwise those for *, with the most specific Allow/Disallow match winning. If robots.txt cannot be read (other than a 404), we do not crawl the site.
RateOne request at a time per site, at least 10 seconds apart, or your Crawl-delay if it is longer. That applies to every page, including about and contact pages.
Content signalsWe pass pages to an AI model to extract contact details, which counts as ai-input. If your robots.txt sets Content-Signal: ai-input=no, we do not crawl your site.
Protected addressesWe do not decode obfuscated email addresses. If your site hides addresses with Cloudflare's email protection or a similar tool, we treat that as you asking not to be harvested, and we do not record the hidden address. We collect only addresses printed in plain text.
RenderingSome pages are rendered through Cloudflare Browser Run. Those requests also carry Cloudflare's signed Signature-Agent headers, which identify the traffic as automated.

How to block or limit it

Add this to your robots.txt to block NewzWyreBot entirely:

User-agent: NewzWyreBot
Disallow: /

Or use the form below to opt your whole domain out. That is recorded on our blocklist, which the crawler checks before every fetch. You can also opt out a single email address.

An email or domain opt-out takes effect immediately and automatically. Access, deletion and correction requests are answered by a person within one month.