Fused Attention Modules Embedded in Artificial Neural Networks for Low Dose CT Denoising With Integrated Loss Functions
Bibliographic record
Abstract
<p>X-ray Computed Tomography (CT) is a non-invasive medical diagnostic tool that has raised public concerns due to the associated health risks of radiation dose to patients. Reducing the radiation dose leads to noise artifacts, making the low-dose CT images unreliable for diagnosis. Hence, low-dose computed tomography (LDCT) image reconstruction techniques have offered a new challenge in the research area. This thesis focuses on reconstructing LDCT images using deep learning techniques to provide an efficient, effective, and accurate training regimes for LDCT image denoising. A fusion of spatial and channel attention modules integrated into a dilated residual network is proposed to improve the structural details of denoised LDCT images. Further, a combination of perceptual loss, per-pixel loss, and structural dissimilarity loss is used for the optimization of the overall network. These objective functions aim to preserve structural details, avoid edge over-smoothing and enhance the image texture, respectively. Peak Signal-to-Noise Ratio (PSNR) and Structural Similarity Index Metrics (SSIM) are used for measuring the quantitative results. A comparative experiment was done between the proposed model and the recent denoising models, such as Block Matching and 3D Filtering (BM3D), patch Markovnian Generative Adversarial Network (patch-GAN) and dilated residual learning with edge detection (DRL-E-MP). Not only with quantitative results, but these models were also compared visually. To further strengthen the validity of the outcomes, five different CT image datasets were used. The proposed model obtained the highest PSNR/SSIM value of 34.36/0.6971 while BM3D resulted in the lowest value with 30.24/0.4461 using the chest dataset from the Mayo Clinic. Overall, the proposed network demonstrated that it could outperform state-of-the-art models.</p>
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.001 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.000 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".