Methodology & Data Sources
PlainEmissions presents the EU EDGAR v8.0 greenhouse-gas dataset as plain, comparable per-country and per-sector pages. This page documents every step of how the raw EDGAR workbooks become the figures you see, the source, sector-taxonomy harmonization, unit conversions, vintage tracking, and limitations, and how three other authoritative datasets (World Bank, Climate TRACE, UNFCCC) differ from EDGAR.
Download the country-level EDGAR extract this methodology describes: country-emissions-statistics.csv.
Data sources
PlainEmissions currently ingests one primary dataset, EU EDGAR, into its database. Every number rendered on a country, sector, or ranking page is an EDGAR figure. Three other authoritative datasets, World Bank Climate Data, Climate TRACE, and UNFCCC national inventories, are documented below and referenced throughout our research and guide pages for methodological context and contrast, but their record-level data is not currently loaded into our database, so no page on this site displays a World Bank, Climate TRACE, or UNFCCC number directly.
1. EU EDGAR (Joint Research Centre) — ingested, the source of every figure on this site
The Emissions Database for Global Atmospheric Research, maintained by the European Commission's Joint Research Centre, is a bottom-up model that estimates greenhouse-gas emissions for every country and major sector from 1970 to the most recent reporting year (typically two years behind real time). License: CC BY 4.0. Source: edgar.jrc.ec.europa.eu. We treat EDGAR as the canonical cross-country comparability layer because the same methodology is applied to every country, politics and reporting capacity do not enter.
2. World Bank Climate Knowledge Portal — referenced, not currently ingested
The World Bank's Climate Change Knowledge Portal aggregates country-level climate and emissions indicators with broad historical coverage and high update frequency. License: CC BY 4.0. Source: climateknowledgeportal.worldbank.org. We describe the World Bank layer here for context (it is the dataset researchers would use for macro indicators, national totals, per-capita normalization, GDP-intensity), but its records are not loaded into our database and no figure on this site is drawn from it.
3. Climate TRACE — referenced, not currently ingested
Climate TRACE is an independent coalition that produces emissions estimates using satellite observations and machine learning at facility-level resolution. License: CC BY 4.0. Source: climatetrace.org. Climate TRACE is the only one of these datasets that does not rely on self-reported inputs, which is why it is the strongest independent check on national inventories in principle, but its records are not loaded into our database and no figure on this site is drawn from it.
4. UNFCCC National Inventory Submissions — referenced, not currently ingested
Annex I and (since 2024) all parties to the UN Framework Convention on Climate Change submit national greenhouse-gas inventories using IPCC reporting guidelines. License: public. Source: unfccc.int. UNFCCC inventories are the official legal record under international climate-treaty obligations, but they are self-reported and reporting capacity varies dramatically across countries; its records are not loaded into our database and no figure on this site is drawn from it.
Supporting references
For cross-checks and complementary indicators we also consult the U.S. EPA Climate Change Indicators (US-specific sectoral validation), World Bank total greenhouse-gas emissions indicator, OECD environmental indicators, IEA energy and emissions statistics, and the open-data archive at Wikipedia's list of countries by greenhouse-gas emissions (as a cross-source sanity check, not a primary source).
Harmonization steps
EDGAR's raw workbooks use their own country codes, sector taxonomies, gas categorizations, and units. The ETL pipeline performs the following normalization steps in order:
- Country codes - every record is mapped to ISO 3166-1 alpha-3. Historical entities
(USSR, Yugoslavia, etc.) are mapped to a controlled successor-state set documented in the
database
countriestable. - Sector taxonomy - upstream categories are mapped to a common IPCC-aligned hierarchy: Energy, Transport, Buildings, Industry, Agriculture, Land use & forestry (LULUCF), Waste, Fugitive emissions. Where upstream sub-sectors do not cleanly map, the record is retained at the lowest unambiguous parent.
- Gas normalization - gases are tracked individually (CO2, CH4, N2O, HFC, PFC, SF6, NF3) and also expressed as CO2-equivalent using IPCC AR6 100-year global-warming potential multipliers. The native unit value is always retained alongside the CO2e value for transparency.
- Provenance - every fact-table row records a
source_code, and the schema is designed to hold WB_CLIMATE, CLIMATE_TRACE, and UNFCCC rows alongside EDGAR's for future multi-source comparison. Today onlyEDGARrows are populated in the live database. - No interpolation - when a country-year-sector-gas combination is missing from EDGAR, it is left missing. We do not impute and we do not back-fill from other sources.
Vintage tracking
Every EDGAR record carries its upstream release vintage. The data_sources table
registers the other three datasets for future reference, but only EDGAR's row currently
reflects an in-database vintage. When EDGAR issues a new release, only its rows are updated;
historic vintages are not retroactively rewritten in published research pages.
Update cadence
Only EDGAR is currently loaded into our database and refreshed. The other three datasets' own publication cadences (for context, should we load them in the future):
- EDGAR (ingested): annual, typically September-November
- World Bank (referenced only): monthly indicator refresh
- Climate TRACE (referenced only): quarterly
- UNFCCC (referenced only): rolling, country-by-country
PlainEmissions refreshes within four weeks of a major EDGAR release. Minor revisions (single-country corrections) propagate via the corrections-overlay framework documented in our about page.
Limitations
- Greenhouse-gas measurement is inherently uncertain. Even the best satellite estimates have error bands of single-digit to double-digit percent at the country level for some sectors. We surface this uncertainty as multi-source spread rather than hiding it inside a single number.
- LULUCF (land use, land-use change, and forestry) is the most-disputed sector across all four sources. Country pages render LULUCF separately and note when source disagreement exceeds 50%.
- UNFCCC inventory coverage is incomplete for some developing countries; for those countries the comparable EDGAR or Climate TRACE figure is the most useful reference.
- Climate TRACE's facility-level estimates are most accurate for large point sources (power plants, cement, steel, refineries) and less accurate for diffuse sources (agriculture, transport).
How figures are sourced
Country and sector data pages are loaded directly from the database and rendered server-side - numbers are never modified between source row and page. Editorial research pages and methodology notes are grounded in the same upstream datasets, EU EDGAR, the World Bank, Climate TRACE, and UNFCCC national inventory submissions. Every research page cites the specific upstream figures it references so claims remain verifiable.
Corpus placement (#N of M)
Country and sector pages show where each entity sits in the PlainEmissions EDGAR corpus, not just the entity's own totals. Placement uses two ladders that can diverge, so a single "#N of M" never pretends one metric is the whole story.
- Countries - absolute total: rank among economies with live EDGAR rows for the latest data year, ordered by summed MtCO2e across all IPCC sectors and gases (largest absolute footprint = #1). China typically leads this ladder.
- Countries - per capita: rank among economies with population > 0, ordered by tCO2e per person (highest intensity = #1). Small high-intensity economies such as the UAE can lead here while sitting mid-pack on absolute totals.
- Sectors - absolute total: rank among IPCC sectors with live EDGAR rows, ordered by summed MtCO2e in the latest year (largest sector footprint = #1).
- Sectors - methane share: rank by CH4's share of each sector's CO2e total (highest methane fraction = #1). Waste and fugitive sectors can lead this ladder while energy leads absolute Mt. Country-count coverage is flat across sectors here, so it is not used as a second ladder.
Ties break by entity name (A-Z). Ranks are recomputed from the live fact table on each request (cached in-process); they are inventory placement inside this corpus, not a climate-policy score or a claim about UNFCCC legal inventories.
Emissions Performance Score
Every country page carries a 0-100 + A-F composite score, built from four percentile- benchmarked dimensions: per-capita emissions, GDP intensity, methane share of the gas mix, and the multi-year emissions trend (first to latest EDGAR-tracked year). Each dimension is scored by linear interpolation against the corpus's own p10/p25/p50/p75/p90 breakpoints below (lower is scored better for all four dimensions), then combined by weight. A dimension with no usable data (e.g. no GDP figure) is dropped and its weight redistributes across the remaining dimensions, so a missing input never silently deflates the score.
- Per-capita emissions (weight 0.30): EDGAR latest-year total ÷ population.
- GDP intensity (weight 0.30): EDGAR latest-year total ÷ GDP.
- Methane share (weight 0.15): CH4's share of the latest-year gas mix.
- Emissions trend (weight 0.25): % change from the first to the latest EDGAR-tracked year for that country.
| Dimension | p10 | p25 | p50 (median) | p75 | p90 |
|---|---|---|---|---|---|
| Per-capita emissions (t/person) | 2.89 | 5.47 | 7.15 | 11.33 | 17.77 |
| GDP intensity (t/$1K GDP) | 0.13 | 0.22 | 0.53 | 0.90 | 1.34 |
| Methane share (%) | 8.8 | 12.9 | 16.9 | 27.8 | 42.3 |
| Emissions trend (% change) | -47.7 | 4.0 | 119.5 | 425.5 | 562.4 |
Breakpoints measured 2026-09-10 across the {51 EDGAR-tracked countries with a valid latest-year total} (the same live corpus the rank ladders above use), not an external or invented benchmark. Every input traces to EDGAR sector/gas rows already sourced elsewhere on this page; the score adds no new data source of its own.
Last updated: