Reviewing the scope and thematic focus of 100 000 publications on energy consumption, services and social aspects of climate change: a big data approach to demand-side mitigation <sup>*</sup>
Bibliographic record
Abstract
Abstract As current action remains insufficient to meet the goals of the Paris agreement let alone to stabilize the climate, there is increasing hope that solutions related to demand, services and social aspects of climate change mitigation can close the gap. However, given these topics are not investigated by a single epistemic community, the literature base underpinning the associated research continues to be undefined. Here, we aim to delineate a plausible body of literature capturing a comprehensive spectrum of demand, services and social aspects of climate change mitigation. As method we use a novel double-stacked expert—machine learning research architecture and expert evaluation to develop a typology and map key messages relevant for climate change mitigation within this body of literature. First, relying on the official key words provided to the Intergovernmental Panel on Climate Change by governments (across 17 queries), and on specific investigations of domain experts (27 queries), we identify 121 165 non-unique and 99 065 unique academic publications covering issues relevant for demand-side mitigation. Second, we identify a literature typology with four key clusters: policy, housing, mobility, and food/consumption. Third, we systematically extract key content-based insights finding that the housing literature emphasizes social and collective action, whereas the food/consumption literatures highlight behavioral change, but insights also demonstrate the dynamic relationship between behavioral change and social norms. All clusters point to the possibility of improved public health as a result of demand-side solutions. The centrality of the policy cluster suggests that political actions are what bring the different specific approaches together. Fourth, by mapping the underlying epistemic communities we find that researchers are already highly interconnected, glued together by common interests in sustainability and energy demand. We conclude by outlining avenues for interdisciplinary collaboration, synthetic analysis, community building, and by suggesting next steps for evaluating this body of literature.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.013 | 0.067 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.002 | 0.002 |
| Bibliometrics | 0.067 | 0.099 |
| Science and technology studies | 0.002 | 0.002 |
| Scholarly communication | 0.009 | 0.007 |
| Open science | 0.001 | 0.004 |
| Research integrity | 0.002 | 0.001 |
| Insufficient payload (model declined to judge) | 0.011 | 0.005 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".