MétaCan
Menu
Back to cohort
Record W4313888578 · doi:10.3390/curroncol30010064

Machine Learning Approaches with Textural Features to Calculate Breast Density on Mammography

2023· article· en· W4313888578 on OpenAlexvenueno aff
Mario Sansone, Roberta Fusco, Francesca Grassi, Gianluca Gatta, Maria Paola Belfiore, Francesca Angelone, Carlo Ricciardi, Alfonso Maria Ponsiglione, Francesco Amato, Roberta Galdiero, Roberta Grassi, Vincenza Granata, Roberto Grassi

Bibliographic record

VenueCurrent Oncology · 2023
Typearticle
Languageen
FieldMedicine
TopicDigital Radiography and Breast Imaging
Canadian institutionsnot available
Fundersnot available
KeywordsMedicineArtificial intelligenceMammographyFalse positive paradoxBreast cancerSupport vector machinePopulationGold standard (test)Machine learningDigital mammographyMedical physicsCancerRadiologyComputer scienceInternal medicine

Abstract

fetched live from OpenAlex

BACKGROUND: breast cancer (BC) is the world's most prevalent cancer in the female population, with 2.3 million new cases diagnosed worldwide in 2020. The great efforts made to set screening campaigns, early detection programs, and increasingly targeted treatments led to significant improvement in patients' survival. The Full-Field Digital Mammograph (FFDM) is considered the gold standard method for the early diagnosis of BC. From several previous studies, it has emerged that breast density (BD) is a risk factor in the development of BC, affecting the periodicity of screening plans present today at an international level. OBJECTIVE: in this study, the focus is the development of mammographic image processing techniques that allow the extraction of indicators derived from textural patterns of the mammary parenchyma indicative of BD risk factors. METHODS: a total of 168 patients were enrolled in the internal training and test set while a total of 51 patients were enrolled to compose the external validation cohort. Different Machine Learning (ML) techniques have been employed to classify breasts based on the values of the tissue density. Textural features were extracted only from breast parenchyma with which to train classifiers, thanks to the aid of ML algorithms. RESULTS: the accuracy of different tested classifiers varied between 74.15% and 93.55%. The best results were reached by a Support Vector Machine (accuracy of 93.55% and a percentage of true positives and negatives equal to TPP = 94.44% and TNP = 92.31%). The best accuracy was not influenced by the choice of the features selection approach. Considering the external validation cohort, the SVM, as the best classifier with the 7 features selected by a wrapper method, showed an accuracy of 0.95, a sensitivity of 0.96, and a specificity of 0.90. CONCLUSIONS: our preliminary results showed that the Radiomics analysis and ML approach allow us to objectively identify BD.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.000
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.538
Threshold uncertainty score0.634

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0000.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0010.001
Science and technology studies0.0000.000
Scholarly communication0.0000.000
Open science0.0000.000
Research integrity0.0000.001
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.077
GPT teacher head0.336
Teacher spread0.259 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designObservational
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations25
Published2023
Admission routes1
Has abstractyes

Explore more

Same venueCurrent OncologySame topicDigital Radiography and Breast ImagingFrench-language works237,207