MétaCan
Menu
Back to cohort
Record W4381429499 · doi:10.1016/j.ostima.2023.100101

NEURAL SHAPE MODELS ENCODE BONE SHAPE FEATURES NOT CAPTURED BY STATISTICAL SHAPE MODELS

2023· article· en· W4381429499 on OpenAlexfundno aff
Anthony A. Gatti, Feliks Kogan, Garry E. Gold, Scott L. Delp, Akshay Chaudhari

Bibliographic record

VenueOsteoarthritis Imaging · 2023
Typearticle
Languageen
FieldMedicine
TopicBone health and osteoporosis research
Canadian institutionsnot available
FundersCanadian Institutes of Health ResearchNational Institutes of Health
KeywordsPattern recognition (psychology)Standard deviationArtificial intelligenceMathematicsSagittal planeFemurShape analysis (program analysis)Computer scienceMedicineStatisticsAnatomy

Abstract

fetched live from OpenAlex

The recently proposed B-Score uses statistical shape models (SSM) to represent femur shape as a scalar value similar to the osteoporosis T-score. The B-Score quantifies OA bone shape and is defined as the distance from the mean healthy bone shape (B-Score=0) to the mean OA bone shape, where 1-unit is equal to the standard deviation of the healthy B-Scores [Bowes et al. 2021]. However, SSMs require finding matching points between subjects’ femurs, and learn linear features, potentially limiting their ability to capture physiologic shape. Neural Shape Models (NSM) have been shown to represent object surfaces without requiring matching points between subjects using non-linear neural networks. Here, we use NSMs to reconstruct bone shapes and use these features to encode information about OA. To compare B-Scores learned from a NSM and a SSM. Data from the 24 and 48-month visit of the right knee of 562 participants enrolled in the OAI were included (335 females, mean age 63.5(8.9) years, BMI 30.8(4.8) kg/m2, and KLG counts of 0=35, 1=79, 2=269, 3=167, 4=12). Fig 1 depicts the data analysis pipeline; sagittal DESS MRIs were segmented using a CNN and femur surfaces were extracted using marching cubes. The NSM and SSM models were fit to the 24-month data of half the subjects. The NSM and SSM learned feature spaces were 256 and 90 dimensions, respectively. Fitted models were used to obtain shape features from the 48-month data of all subjects. Finally, NSM and SSM B-scores were computed to assess how the NSM and SSM feature spaces affect the learned B-scores. To determine whether each model has the capacity to represent the other's B-Score, the amount of variance in the B-Score explained by the feature space of the other model was calculated using linear regression. Since the B-Score produces a range of scores within each KL grade, the distribution of B-Scores per KLG were plotted. The odds ratios (OR) for knee pain and TKA were computed between B-Score quartiles (1 vs 2, 3, 4) in OA knees (KLG >=2). Pain was defined using previous criteria [Morales et al. 2021]. The NSM explained 82% of the variance in SSM B-score, yet the SSM only explained 55% of the variance in NSM B-score (Fig 2). Fig 3 shows the distribution of B-Scores per KLG demonstrating that within a KLG there is a range of B-Scores providing more specific shape information. Table 1 includes ORs for pain and TKA between quartile 1 and all other quartiles for both B-scores. The NSM learned non-linear shape features that encode clinically relevant information about OA without a need to find matching points between subjects. The NSM and SSM performed similarly for predicting clinical outcomes. The NSM had a broader range of B-scores between KLGs primarily driven by a large difference between KLG 0 and 1, potentially indicating greater expressivity for the NSM. The SSM did a poor job predicting the NSM B-score, which was particularly evident in KLG 0 knees where the SSM was unable to reproduce NSM B-Scores in the healthy range (Fig 2). Small samples of KLG 0 and 4 knees likely limit both SSM and NSM B-Scores. More data from the OAI will likely enable the more flexible NSM to learn more expressive representations particularly in under-represented sub-samples, like KLG 4 knees. Results from this study indicate that the NSM captures novel bone shape information that cannot be learned by the SSM.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.001
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesMeta-epidemiology (narrow)
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Simulation or modeling · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.979
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0010.000
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.000
Bibliometrics0.0000.001
Science and technology studies0.0000.000
Scholarly communication0.0000.001
Open science0.0000.000
Research integrity0.0000.001
Insufficient payload (model declined to judge)0.0010.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.031
GPT teacher head0.314
Teacher spread0.283 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designSimulation or modeling
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2023
Admission routes1
Has abstractyes

Explore more

Same venueOsteoarthritis ImagingSame topicBone health and osteoporosis researchFrench-language works237,207