Abstract 2607: The cBioPortal for Cancer Genomics: an open source platform for accessing and interpreting complex cancer genomics data in the era of precision medicine
Notice bibliographique
Résumé
Abstract The cBioPortal for Cancer Genomics is an open-access portal (http://cbioportal.org) that enables interactive, exploratory analysis of large-scale cancer genomics data. It integrates genomic and clinical data, and provides a suite of visualization and analysis options, including cohort and patient-level visualization, mutation visualization, survival analysis, enrichment analysis, and network analysis. The user interface is user-friendly, responsive, and makes genomic data easily accessible to translational scientists, biologists, and clinicians. The cBioPortal is a fully open source platform. All code is available on GitHub (https://github.com/cBioPortal/) under GNU Affero GPL license. The code base is maintained by multiple groups, including Memorial Sloan Kettering Cancer Center, Dana-Farber Cancer Institute, Children’s Hospital of Philadelphia, Princess Margaret Cancer Centre, and The Hyve, an open source bioinformatics company based in the Netherlands. More than 30 academic centers as well as multiple pharmaceutical and biotech companies maintain private instances of the cBioPortal. This includes the recently launched cBioPortal instance at the NCI Genomic Data Commons (https://cbioportal.gdc.nci.nih.gov/), and two large cBioPortal instances hosting genomic and clinical data at MSK and DFCI, supporting the MSK-IMPACT and DFCI Profile projects, two of the largest clinical sequencing efforts in the world. Our multi-institutional software team has accelerated the progress of evolving the core architectural technologies and developing new features to keep pace with the rapidly advancing fields of cancer genomics and precision cancer medicine. For example, we have integrated multi-platform genomics data with extensive clinical data including patient demographics, treatment history, and survival data. We have also developed a patient-centric view that visualizes both clinical and genomic data with annotation from OncoKB knowledge base. In the next few years, the development team will focus on the following areas: (1) Implementing major architectural changes to ensure future scalability and performance. (2) New features to support precision medicine, including (i) improved integration of knowledge base annotation, (ii) enhanced visualization of patient timeline, drug response, and tumor evolution, (iii) new patient similarity metrics, (iv) improved support for immunogenomics and immunotherapy, and (v) new visualization and analysis features for understanding response to therapy. (3) New analysis and target discovery features for large cohorts, including (i) supporting user-defined virtual cohort by selecting samples from multiple studies, and (ii) comparison of genomic or clinical characteristics of two or more selected cohorts. (4) Expanding community outreach, user support and training, and documentation. Citation Format: Jianjiong Gao, Ersin Ciftci, Pichai Raman, Pieter Lukasse, Istemi Bahceci, Adam Abeshouse, Hsiao-Wei Chen, Ino de Bruijn, Benjamin Gross, Zachary Heins, Ritika Kundra, Aaron Lisman, Angelica Ochoa, Robert Sheridan, Onur Sumer, Yichao Sun, Jiaojiao Wang, Manda Wilson, Hongxin Zhang, James Xu, Andy Dufilie, Priti Kumari, James Lindsay, Anthony Cros, Karthik Kalletla, Fedde Schaeffer, Sander Tan, Sjoerd van Hagen, Jorge Reis-Filho, Kees van Bochove, Ugur Dogrusoz, Trevor Pugh, Adam Resnick, Chris Sander, Ethan Cerami, Nikolaus Schultz. The cBioPortal for Cancer Genomics: an open source platform for accessing and interpreting complex cancer genomics data in the era of precision medicine [abstract]. In: Proceedings of the American Association for Cancer Research Annual Meeting 2017; 2017 Apr 1-5; Washington, DC. Philadelphia (PA): AACR; Cancer Res 2017;77(13 Suppl):Abstract nr 2607. doi:10.1158/1538-7445.AM2017-2607
Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.
Comment cette classification a été obtenuedéplier
Prédiction machine sur la base complète
Imitation des enseignantsNi prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.
Scores du classifieur distillé par catégorie (deux têtes)
| Catégorie | Codex | Gemma |
|---|---|---|
| Métarecherche | 0,005 | 0,015 |
| Méta-épidémiologie (sens strict) | 0,002 | 0,002 |
| Méta-épidémiologie (sens large) | 0,002 | 0,002 |
| Bibliométrie | 0,006 | 0,006 |
| Études des sciences et des technologies | 0,001 | 0,001 |
| Communication savante | 0,005 | 0,004 |
| Science ouverte | 0,006 | 0,010 |
| Intégrité de la recherche | 0,003 | 0,005 |
| Charge utile insuffisante (le modèle a refusé de juger) | 0,064 | 0,078 |
Scores machine (provisoires)
Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.
Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.
score_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découleClassification
machine, non validéePrédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.
Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».