BloggedBot
The crawler that reads public pages on your behalf, what it accesses, and how a site owner can ask to be excluded.
BloggedBot is the crawler Blogged uses to read public web pages. It identifies itself in every request:
BloggedBot/0.1 (+https://blogged.dev/bot)What it reads and why
- Your own site, during setup
To establish what your product is, how it is positioned, and what it looks like. See Sources.
- Approved sources, on refresh
To keep the facts behind your writing current. See refreshing knowledge.
- Approved competitor sites
To support comparison and Competitor Blog Watch.
What it does not do
It reads pages a visitor could open. It does not sign in, submit forms, or reach anything behind authentication, and it does not collect personal data from the pages it reads.
Limits
Reads are bounded: page budgets per ingest, a response size ceiling, a request timeout, and cooldowns between runs against the same site. Competitor checks are rate limited separately from full runs.
Asking to be excluded
Site owners who would rather not be read can write to hello@blogged.dev. There is more detail, including the public statement of what the crawler is for, at blogged.dev/bot.