The data
Everything behind this site is free to download and reuse. If you publish something with it, a link back is appreciated.
The files are hosted in the project’s GitHub repository, alongside the site lists and the crawler, so you can also clone it or follow its history.
This week, as a spreadsheet
One row per site: body font, heading font, platform, how the fonts are served, and when we saw it.
Weekly crawls
Every tracked homepage, once a week. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.
- 2026-W393.1 MB
Daily crawls
A few hundred of the best-known sites, every day. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.
- 2026-09-2384 KB
Archive backfill
Quarterly estimates from the Internet Archive and Arquivo.pt. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.