Expanding TheCellVision.org: a central repository for visualizing and mining high-content cell imaging projects
Bibliographic record
Abstract
We previously constructed TheCellVision.org, a central repository for visualizing and mining data from yeast high-content imaging projects. At its inception, TheCellVision.org housed two high-content screening (HCS) projects providing genome-scale protein abundance and localization information for the budding yeast Saccharomyces cerevisiae, as well as a comprehensive analysis of the morphology of its endocytic compartments upon systematic genetic perturbation of each yeast gene. Here, we report on the expansion of TheCellVision.org by the addition of two new HCS projects and the incorporation of new global functionalities. Specifically, TheCellVision.org now hosts images from the Cell Cycle Omics project, which describes genome-scale cell cycle-resolved dynamics in protein localization, protein concentration, gene expression, and translational efficiency in budding yeast. Moreover, it hosts PIFiA, a computational tool for image-based predictions of protein functional annotations. Across all its projects, TheCellVision.org now houses >800,000 microscopy images along with computational tools for exploring both the images and their associated datasets. Together with the newly added global functionalities, which include the ability to query genes in any of the hosted projects using either yeast or human gene names, TheCellVision.org provides an expanding resource for single-cell eukaryotic biology.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".