Longitudinal functional changes with clinically significant radiographic progression in idiopathic pulmonary fibrosis: are we following the right parameters?
Bibliographic record
Abstract
BACKGROUND: Progression of the disease in idiopathic pulmonary fibrosis (IPF) is difficult to predict, due to its variable and heterogenous course. The relationship between radiographic progression and functional decline in IPF is unclear. We sought to confirm that a simple HRCT fibrosis visual score is a reliable predictor of mortality in IPF, when longitudinally followed; and to ascertain which pulmonary functional variables best reflect clinically significant radiographic progression. METHODS: One-hundred-twenty-three consecutive patients with IPF from 2 centers were followed for an average of 3 years. Longitudinal changes of HRCT fibrosis scores, forced vital capacity (FVC), total lung capacity and diffusing lung capacity for carbon monoxide were considered. HRCTs were scored by 2 chest radiologists. The primary outcome was lung transplant (LTx)-free survival after the follow-up HRCT. RESULTS: During the follow-up period, 43 deaths and 11 LTx occurred. On average, the HRCT fibrosis score increased significantly, and a longitudinal increase > 7% predicted LTx-free survival significantly, with good specificity, but limited sensitivity. The correlation between radiographic and functional progression was moderately significant. HRCT progression and FVC decline predicted LTx-free survival independently and significantly, with better sensitivity, but worse specificity for a ≥ 5% decline of FVC. However, the area under the curve towards LTx-survival were only 0.61 and 0.62, respectively. CONCLUSIONS: The HRCT fibrosis visual score is a reliable and responsive tool to detect clinically meaningful disease progression. Although no individual pulmonary function test closely reflects radiographic progression, a longitudinal FVC decline improves sensitivity in the detection of clinically significant disease progression. However, the accuracy of these methods remains limited, and better prognostication models need to be found.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.003 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.002 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.002 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".