The new ACR/EULAR criteria for rheumatoid arthritis can identify patients with same disease activity but less damage by ultrasound
Bibliographic record
Abstract
OBJECTIVE: We aimed to compare the ultrasound findings of patients fulfilling the 1987 ACR [OLD-rheumatoid arthritis (RA)] and the new ACR/EULAR (NEW-RA) classification criteria to examine the impact of the new criteria on disease characteristics, particularly disease duration. MATERIAL AND METHODS: A total of 2730 hands, wrists, elbows, knees, ankles, and foot joints of 105 consecutive patients with inflammatory arthritis, i.e., 82 patients fulfilling the RA criteria (60 patients, OLD-RA; 22 patients, NEW-RA alone) and 23 patients with undifferentiated arthritis, were scanned using ultrasound. Synovitis, erosions, and power Doppler (PD) findings were scored using a scale of 0-3 and scores form each joint were added up to calculate synovitis, PD and erosion scores for each patient. RESULTS: OLD-RA and NEW-RA patients had similar swollen joint count, tender joint count, acute-phase response, patient global, and disease activity assessment 28 scores. The disease duration was longer in OLD-RA patients [30 (3-179) months] than in NEW-RA patients [16 (0-45) months; p=0.009]. Both the groups had similar synovitis and PD scores, whereas erosion scores were higher in OLD-RA patients than in NEW-RA patients (p=0.009). Patients with undifferentiated arthritis were older than those with RA and had fewer swollen joints than NEW-RA patients [0 (0-4) vs. 2 (0-9); p=0.017]. All other disease activity parameters were similar in both NEW-RA and OLD-RA patients. Both the synovitis (p=0.006) and erosion (p=0.007) scores were lower in patients with undifferentiated arthritis than in OLD-RA patients, despite the scores being similar to those in NEW-RA patients. CONCLUSION: The new ACR/EULAR RA criteria enabled the classification of patients with similar disease activity (by clinical assessment and ultrasound) but with less damage. A similar disease activity should ensure suitability for an intervention, and a shorter duration and less damage should improve the outcome with patient benefit.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.002 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.001 | 0.001 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".