webtracking.org · open data

Reference data for the tracking ecosystem

The entity map, cookie database, and AI crawler/referrer registries behind The State of Web Tracking — versioned JSON, explorable below, growing every quarter through curation queues generated from the crawl itself.

CC BY 4.0 — commercial use OK, attribute webtracking.org measured against 15.7M pages updated —

Use it as an API

Stable versioned endpoints, CORS-enabled. latest moves; snapshots never do.

GET https://data.webtracking.org/v1/latest/entities.json GET https://data.webtracking.org/v1/2026Q3/entities.json GET https://data.webtracking.org/v1/latest/queues/unmapped-domains.json GET https://data.webtracking.org/index.json

Wrong owner? Missing cookie? Open a PR — the curation queues are the ranked to-do list. Corrections are applied, dated, and credited. · Crawl aggregates: index.json lists every CSV.