Testing of the preliminary OMERACT validation criteria for a biomarker to be regarded as reflecting structural damage endpoints in rheumatoid arthritis clinical trials: the example of C-reactive protein.
Bibliographic record
Abstract
OBJECTIVE: A list of 14 criteria for guiding the validation of a soluble biomarker as reflecting structural damage endpoints in rheumatoid arthritis (RA) clinical trials was drafted by an international working group after a Delphi consensus exercise. C-reactive protein (CRP), a soluble biomarker extensively studied in RA, was then used to test these criteria. Our objectives were: (1) To assess the strength of evidence in support of CRP as a soluble biomarker reflecting structural damage in RA according to the draft validation criteria. (2) To assess the strength of recommendation for inclusion of individual criteria in the draft set. METHODS: A systematic literature review was conducted to elicit evidence in support of each specific criterion composing the 14-criteria draft set. A summary of the key literature findings per criterion was presented to both the working group and to participants in a special interest soluble biomarker group at OMERACT 8. Participants at OMERACT 8 were asked to rate the strength of evidence and the strength of the recommendation in support of each individual criterion on a 0-10 numerical rating scale. Working group members not present at OMERACT voted by a Web-based survey. RESULTS: Minimal data were extracted from the literature pertaining to those criteria listed under the category of truth. Ratings for strength of evidence were moderate to low (< 7) for CRP as a biomarker reflecting structural damage in RA; this was true for all criteria except those listed under the category of feasibility and 2 listed under the category of discrimination pertaining to assay reproducibility and evidence regarding sources of variability. Ratings for strength of recommendation for inclusion of each of the 14 criteria in the draft set were high (> 7) except for those criteria listed under the category of truth. CONCLUSION: The draft criteria serve as a useful template in the evaluation of the strength of evidence in support of a particular soluble biomarker as reflecting structural damage in RA.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.515 | 0.737 |
| Meta-epidemiology (narrow) | 0.003 | 0.003 |
| Meta-epidemiology (broad) | 0.006 | 0.018 |
| Bibliometrics | 0.011 | 0.007 |
| Science and technology studies | 0.004 | 0.005 |
| Scholarly communication | 0.008 | 0.007 |
| Open science | 0.008 | 0.011 |
| Research integrity | 0.014 | 0.010 |
| Insufficient payload (model declined to judge) | 0.003 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; the direct Gemma label and the distilled Codex classifier agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".