Bibliographic record
Abstract
Within the framework of a common monetary policy, the monitoring of economic developments in the euro area, on a regular basis, is of particular importance. Despite of the ongoing improvement, the data available for the euro area as a whole are still relatively limited and released with some lag. The assessment of the economic situation requires synthetic measures representative of activity in the economy as a whole. Gross Domestic Product (GDP) is the best measure acknowledged for this purpose. However, GDP is only made available on a quarterly basis and released with a significant lag, which makes it difficult to assess economic activity on a regular and timely basis. In fact, the first estimate for the euro area GDP in a given quarter is released 70 days after the end of that quarter.(1) Thus, one needs to resort to other synthetic measures which provide information on economic developments in the euro area on a more timely and frequent basis. The purpose of this article is to evaluate the performance of several economic composite indicators, which are currently released on a regular basis by several institutions, including the European Commission, the Organisation for Economic Co-operation and Development (OECD) and the Centre for Economic Policy Research (CEPR). The aim of this article is to assess to what extent these composite indicators allow the monitoring of GDP growth. For this purpose, we resort both to time and frequency domain analysis. This article is organised as follows. Section 2 makes a brief description of the methodology used to evaluate the composite indicators. Section 3 presents the main features of the indicators released by the different institutions and makes an overall assessment of their performance. Section 4 addresses other issues regarding the practical use of the indicators and section 5 concludes.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.004 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".