A lysosomal enigma CLN5 and its significance in understanding neuronal ceroid lipofuscinosis
Bibliographic record
Abstract
Neuronal Ceroid Lipofuscinosis (NCL), also known as Batten disease, is an incurable childhood brain disease. The thirteen forms of NCL are caused by mutations in thirteen CLN genes. Mutations in one CLN gene, CLN5, cause variant late-infantile NCL, with an age of onset between 4 and 7 years. The CLN5 protein is ubiquitously expressed in the majority of tissues studied and in the brain, CLN5 shows both neuronal and glial cell expression. Mutations in CLN5 are associated with the accumulation of autofluorescent storage material in lysosomes, the recycling units of the cell, in the brain and peripheral tissues. CLN5 resides in the lysosome and its function is still elusive. Initial studies suggested CLN5 was a transmembrane protein, which was later revealed to be processed into a soluble form. Multiple glycosylation sites have been reported, which may dictate its localisation and function. CLN5 interacts with several CLN proteins, and other lysosomal proteins, making it an important candidate to understand lysosomal biology. The existing knowledge on CLN5 biology stems from studies using several model organisms, including mice, sheep, cattle, dogs, social amoeba and cell cultures. Each model organism has its advantages and limitations, making it crucial to adopt a combinatorial approach, using both human cells and model organisms, to understand CLN5 pathologies and design drug therapies. In this comprehensive review, we have summarised and critiqued existing literature on CLN5 and have discussed the missing pieces of the puzzle that need to be addressed to develop an efficient therapy for CLN5 Batten disease.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.001 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.001 | 0.002 |
| Open science | 0.001 | 0.001 |
| Research integrity | 0.001 | 0.002 |
| Insufficient payload (model declined to judge) | 0.003 | 0.002 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".