The antinuclear antibody HEp-2 indirect immunofluorescence assay: a survey of laboratory performance, pattern recognition and interpretation
Bibliographic record
Abstract
BACKGROUND: To evaluate the interpretation and reporting of antinuclear antibodies (ANA) by indirect immunofluorescence assay (IFA) using HEp-2 substrates based on common practice and guidance by the International Consensus on ANA patterns (ICAP). METHOD: Participants included two groups [16 clinical laboratories (CL) and 8 in vitro diagnostic manufacturers (IVD)] recruited via an email sent to the Association of Medical Laboratory Immunologists (AMLI) membership. Twelve (n = 12) pre-qualified specimens were distributed to participants for testing, interpretation and reporting HEp-2 IFA. Results obtained were analyzed for accuracy with the intended and consensus response for three main categorical patterns (nuclear, cytoplasmic and mitotic), common patterns and ICAP report nomenclatures. The distributions of antibody titers of specimens were also compared. RESULTS: Laboratories differed in the categorical patterns reported; 8 reporting all patterns, 3 reporting only nuclear patterns and 5 reporting nuclear patterns with various combinations of other patterns. For all participants, accuracy with the intended response for the categorical nuclear pattern was excellent at 99% [95% confidence interval (CI): 97-100%] compared to 78% [95% CI 67-88%] for the cytoplasmic, and 93% [95% CI 86%-100%] for mitotic patterns. The accuracy was 13% greater for the common nomenclature [87%, 95% CI 82-90%] compared to the ICAP nomenclature [74%, 95% CI 68-79%] for all participants. Participants reporting all three main categories demonstrated better performances compared to those reporting 2 or less categorical patterns. The average accuracies varied between participant groups, however, with the lowest and most variable performances for cytoplasmic pattern specimens. The reported titers for all specimens varied, with the least variability for nuclear patterns and most titer variability associated with cytoplasmic patterns. CONCLUSIONS: Our study demonstrated significant accuracy for all participants in identifying the categorical nuclear staining as well as traditional pattern assignments for nuclear patterns. However, there was less consistency in reporting cytoplasmic and mitotic patterns, with implications for assigning competencies and training for clinical laboratory personnel.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.001 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".