Machine Learning Approaches with Textural Features to Calculate Breast Density on Mammography
Bibliographic record
Abstract
BACKGROUND: breast cancer (BC) is the world's most prevalent cancer in the female population, with 2.3 million new cases diagnosed worldwide in 2020. The great efforts made to set screening campaigns, early detection programs, and increasingly targeted treatments led to significant improvement in patients' survival. The Full-Field Digital Mammograph (FFDM) is considered the gold standard method for the early diagnosis of BC. From several previous studies, it has emerged that breast density (BD) is a risk factor in the development of BC, affecting the periodicity of screening plans present today at an international level. OBJECTIVE: in this study, the focus is the development of mammographic image processing techniques that allow the extraction of indicators derived from textural patterns of the mammary parenchyma indicative of BD risk factors. METHODS: a total of 168 patients were enrolled in the internal training and test set while a total of 51 patients were enrolled to compose the external validation cohort. Different Machine Learning (ML) techniques have been employed to classify breasts based on the values of the tissue density. Textural features were extracted only from breast parenchyma with which to train classifiers, thanks to the aid of ML algorithms. RESULTS: the accuracy of different tested classifiers varied between 74.15% and 93.55%. The best results were reached by a Support Vector Machine (accuracy of 93.55% and a percentage of true positives and negatives equal to TPP = 94.44% and TNP = 92.31%). The best accuracy was not influenced by the choice of the features selection approach. Considering the external validation cohort, the SVM, as the best classifier with the 7 features selected by a wrapper method, showed an accuracy of 0.95, a sensitivity of 0.96, and a specificity of 0.90. CONCLUSIONS: our preliminary results showed that the Radiomics analysis and ML approach allow us to objectively identify BD.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".