Distinct isoform of FABP7 revealed by screening for retroelement-activated genes in diffuse large B-cell lymphoma
Bibliographic record
Abstract
Remnants of ancient transposable elements (TEs) are abundant in mammalian genomes. These sequences harbor multiple regulatory motifs and hence are capable of influencing expression of host genes. In response to environmental changes, TEs are known to be released from epigenetic repression and to become transcriptionally active. Such activation could also lead to lineage-inappropriate activation of oncogenes, as one study described in Hodgkin lymphoma. However, little further evidence for this mechanism in other cancers has been reported. Here, we reanalyzed whole transcriptome data from a large cohort of patients with diffuse large B-cell lymphoma (DLBCL) compared with normal B-cell centroblasts to detect genes ectopically expressed through activation of TE promoters. We have identified 98 such TE-gene chimeric transcripts that were exclusively expressed in primary DLBCL cases and confirmed several in DLBCL-derived cell lines. We further characterized a TE-gene chimeric transcript involving a fatty acid-binding protein gene (LTR2-FABP7), normally expressed in brain, that was ectopically expressed in a subset of DLBCL patients through the use of an endogenous retroviral LTR promoter of the LTR2 family. The LTR2-FABP7 chimeric transcript encodes a novel chimeric isoform of the protein with characteristics distinct from native FABP7. In vitro studies reveal a dependency for DLBCL cell line proliferation and growth on LTR2-FABP7 chimeric protein expression. Taken together, these data demonstrate the significance of TEs as regulators of aberrant gene expression in cancer and suggest that LTR2-FABP7 may contribute to the pathogenesis of DLBCL in a subgroup of patients.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".