The data

Everything behind this site is free to download and reuse. If you publish something with it, a link back is appreciated.

The files are hosted in the project’s GitHub repository, alongside the site lists and the crawler, so you can also clone it or follow its history.

This week, as a spreadsheet

One row per site: body font, heading font, platform, how the fonts are served, and when we saw it.

Download CSV

Weekly crawls

Every tracked homepage, once a week. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.

Daily crawls

A few hundred of the best-known sites, every day. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.

Archive backfill

Quarterly estimates from the Internet Archive and Arquivo.pt. Gzipped JSON Lines, one object per site with the full detail: every font on the page and its share of the text, declared faces, font files and bytes.