The language atlas as a file
The atlas is published as one file. Each row is a single language claim, tied to the column it belongs in and to the page it came from. As of 2026-09-12.
Language lists are the fastest-moving claims on this site and the least often dated. A file with a reading date on every row makes it possible to see a list grow rather than only to see its current state.
The column an entry belongs to is carried in the file, so an interface claim can never be counted as an output claim by whoever rebuilds the comparison. That conflation is the single most common error in coverage tables elsewhere.
Rows that record an absence are exported alongside the rest. A market with no claims is the finding a buyer in that market most needs, and dropping those rows would make the file look tidier and mean less.
Reuse is governed by the licence named below. Whatever it requires, the source and date columns are the part worth keeping: a coverage claim without them cannot be re-checked.
1What each column holds
| Column | What it holds |
|---|---|
| fact_id | A stable identifier for the claim, unchanged by later corrections. |
| value | The recorded claim, in the atlas's own wording. |
| applies_to | Which column it fills and which tool or market it concerns. |
| source_url | The vendor page the claim was read from. |
| source_name | That page's own title. |
| value_since | The date the claim is known to hold from. |
| verified | The day the page was last read. |
Inclusion rule. Every recorded claim is exported, including those recording that a vendor names nothing. Order. File order follows recording order, grouping rows by tool.
The file is at /datasets/languages.csv, licensed ODbL 1.0. Cite the atlas and the reading date.
The rest of the atlas: Markets, Tools, Output, Questions, Learn. How a list earns a column is set out on the counting page.