CrawlPact

Registry Observatory

Every figure below is computed directly from CrawlPact's immutable registry release 2026.07.3, resolved from the same frozen release snapshot production evaluation and historical scan rendering use — never from a live, mutable table, and never from the static crawler directory pages.

Data collected
2026-09-07T03:36:52.768Z
Registry release
2026.07.3
Checksum
not computed for this release
Methodology
policy-observatory-methodology-v1

Crawler and operator counts

Crawler and operator summary counts
Tracked crawler records23
Evaluation-eligible (active, deprecated, replaced)23
Distinct operators1

Purpose distribution

Denominator: 23 tracked crawler records.

Crawler records by declared purpose
PurposeRecords
advertising validation2
agent2
mixed1
research1
search7
training5
unknown1
user triggered4

Lifecycle distribution

Crawler records by lifecycle status
Lifecycle statusRecords
active23

Operators by purpose

Which purposes each operator publishes crawlers for
OperatorPurposes
Unknown operatoradvertising validation, agent, mixed, research, search, training, unknown, user triggered

This lists which purposes each operator currently documents separate crawlers for — it is not a transparency or trust ranking (no operator quality score exists).

Source-verification freshness

Denominator: 23 tracked crawler records.

Source-verification freshness summary
Records with a recorded verification date0
Never verified23
Median verification age (days)n/a
Due for re-verification (evaluation-eligible, >180 days or never)23 of 23

Release history

1 published release(s). Each row compares that release to the one immediately before it — evidence-only and editorial-only changes never triggered customer re-evaluation (see methodology).

Registry release history and per-release changes
ReleasePublishedAddedRemovedPurpose changesToken changesEvidence-only
2026.07.3(active)7/28/2026230000

Limitations

  • These figures describe CrawlPact's own crawler-identity registry, not a sample of the web — they do not measure how many websites allow or block any crawler.
  • Purpose classification reflects each operator's published documentation as independently reverified by CrawlPact; it does not measure actual crawler behaviour.
  • An operator may run undocumented or unverified crawlers not yet reflected in this registry.

Analytics preferences

CrawlPact uses optional Google Analytics on public marketing pages to understand which content is useful. It is not used in the authenticated app or admin areas. You can accept or decline analytics without affecting the service.