Profiling gene promoter occupancy of Sox2 in two phenotypically distinct breast cancer cell subsets using chromatin immunoprecipitation and genome-wide promoter microarrays
Bibliographic record
Abstract
INTRODUCTION: Aberrant expression of the embryonic stem cell marker Sox2 has been reported in breast cancer (BC). We previously identified two phenotypically distinct BC cell subsets separated based on their differential response to a Sox2 transcription activity reporter, namely the reporter-unresponsive (RU) and the more tumorigenic reporter-responsive (RR) cells. We hypothesized that Sox2, as a transcription factor, contributes to their phenotypic differences by mediating differential gene expression in these two cell subsets. METHODS: We used chromatin immunoprecipitation and a human genome-wide promoter microarray (ChIP-chip) to determine the promoter occupancies of Sox2 in the MCF7 RU and RR breast cancer cell populations. We validated our findings with conventional chromatin immunoprecipitation, quantitative reverse transcription polymerase chain reaction (qPCR), and western blotting using cell lines, and also performed qPCR using patient RU and RR samples. RESULTS: We found a largely mutually exclusive profile of gene promoters bound by Sox2 between RU and RR cells derived from MCF7 (1830 and 456 genes, respectively, with only 62 overlapping genes). Sox2 was bound to stem cell- and cancer-associated genes in RR cells. Using quantitative RT-PCR, we confirmed that 15 such genes, including PROM1 (CD133), BMI1, GPR49 (LGR5), and MUC15, were expressed significantly higher in RR cells. Using siRNA knockdown or enforced expression of Sox2, we found that Sox2 directly contributes to the higher expression of these genes in RR cells. Mucin-15, a novel Sox2 downstream target in BC, contributes to the mammosphere formation of BC cells. Parallel findings were observed in the RU and RR cells derived from patient samples. CONCLUSIONS: In conclusion, our data supports the model that the Sox2 induces differential gene expression in the two distinct cell subsets in BC, and contributes to their phenotypic differences.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".