Vlad public catalog bundle -- License
=======================================

Creative Commons Attribution 4.0 International (CC BY 4.0)
https://creativecommons.org/licenses/by/4.0/

You are free to share and adapt this material for any purpose, including
commercially, provided you give appropriate credit.

Attribution: Anderson, N. (2026). Vlad. https://ussr.heterodata.org.

WHAT THIS BUNDLE IS
This is the honest table-level CATALOG (per-table metadata) of the Imperial Russian
and Soviet statistical corpus, as ingested by the Robert Database Framework from the
canonical USSR Robert database. It records WHAT WAS EXTRACTED -- titles (English +
Russian original), units, geography, period coverage, taxonomy topic, table
dimensions and two-axis transcription quality per table.

WHAT THIS BUNDLE IS NOT
It contains NO observation (cell) values and NO harmonized economic time series --
those are scoped, not yet constructed.

ENRICHMENT DEPTH IS UNEVEN -- READ THIS BEFORE USING title_en OR topic
The bundle carries 55,749 Tier-A/B tables, but they are not uniformly deep.
25,677 of them (46%) carry the full record: an English
title, a ratified-taxonomy topic and geography facet, and a parsed period. The
remaining 30,072 were metadata-enriched later and have NOT yet been
through the English-translation or taxonomy-classification passes, so for those rows
title_en, topic, geo_facet and start_year are EMPTY while title_raw (Russian),
geography, units and the two-axis quality columns are populated. Empty is honest
absence -- NOT CAPTURED -- never a zero, and never a value to impute. Filter on
topic IS NOT NULL if you need the classified subset.

Snapshot vintage: the Robert-DB slice was baked 2026-08-13 from the canonical USSR
Robert database; document-level campaign figures were reconciled 2026-06-07. This is
a static snapshot with no refresh cadence.

Underlying source statistics are Imperial Russian / Soviet official publications.
