Examination of Pulse Oximetry Tracings to Detect Obstructive Sleep Apnea in Patients with Advanced Chronic Obstructive Pulmonary Disease
Bibliographic record
Abstract
Nocturnal hypoxemia and obstructive sleep apnea (OSA) are common comorbidities in patients with chronic obstructive pulmonary disease (COPD). The authors sought to develop a strategy to interpret nocturnal pulse oximetry and assess its capacity for detection of OSA in patients with stage 3 to stage 4 COPD. A review of consecutive patients with COPD who were clinically prescribed oximetry and polysomnography was conducted. OSA was diagnosed if the polysomnographic apnea-hypopnea index was >15 events⁄h. Comprehensive criteria were developed for interpretation of pulse oximetry tracings through iterative validation and interscorer concordance of ≥80%. Criteria consisted of visually identified desaturation 'events' (sustained desaturation ≥4%, 1 h time scale), 'patterns' (≥3 similar desaturation⁄saturation cycles, 15 min time scale) and the automated oxygen desaturation index. The area under the curve (AUC), sensitivity, specificity and accuracy were calculated. Of 59 patients (27 male), 31 had OSA (53%). The mean forced expiratory volume in 1 s was 46% of predicted (range 21% to 74% of predicted) and 52% of patients were on long-term oxygen therapy. Among 59 patients, 35 were correctly identified as having OSA or not having OSA, corresponding to an accuracy of 59%, with a sensitivity and specificity of 59% and 60%, respectively. The AUC was 0.57 (95% CI 0.55 to 0.59). Using software-computed desaturation events (hypoxemia ≥4% for ≥10 s) indexed at ≥15 events⁄h of sleep as diagnostic criteria, sensitivity was 60%, specificity was 63% and the AUC was 0.64 (95%CI 0.62 to 0.66). No single criterion demonstrated important diagnostic utility. Pulse oximetry tracing interpretation had a modest diagnostic value in identifying OSA in patients with moderate to severe COPD.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.002 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".