MétaCan
Menu
Back to cohort
Record W2041099147 · doi:10.1109/oceans.2010.5664303

Ocean observatories and social computing: Potential and progress

2010· article· en· W2041099147 on OpenAlexaffabout
Dwight Owens, Mairi Best, Eric Guillemot, Reyna Jenkyns, B. Pirenne

Bibliographic record

Venuenot available
Typearticle
Languageen
FieldDecision Sciences
TopicScientific Computing and Data Management
Canadian institutionsEnviro Neptune (Canada)
Fundersnot available
KeywordsComputer scienceWorld Wide WebData scienceOutreachCitizen scienceData managementThe InternetData visualizationVariety (cybernetics)VisualizationObservatoryDatabase

Abstract

fetched live from OpenAlex

In December 2009, after years of planning, preparations and extensive infrastructure deployment, the world's first regional-scale underwater ocean observatory was open for business. NEPTUNE Canada opened its instrument network and data archive to free and open access by anyone willing to register for an account. Thus, we have embarked on a journey to transform our observatory into an online platform for collaborative, multidisciplinary e-science. Four main areas of Internet-mediated activity characterize e-science: data provision, analysis & visualization, collaboration and publication. Data provision entails making our large and ever expanding data archive accessible and searchable through the Web. To support online analysis & visualization, tools must be developed, which allow scientists to display and manipulate a wide variety of data products derived from measurements gathered by the various instruments attached to the observatory. Virtual collaboration can be fostered by making it easy for groups of geographically or institutionally separated researchers to design experiments, control instruments, share analyses and discuss conclusions within a shared web-based workspace. Publication and dissemination of research findings can be supported by tools that help researchers manage and contribute to both informal outreach (e.g. blogs) and the iterative review and revision cycles required for formal manuscript authoring. E-science promises some tantalizing advantages over traditional approaches. By providing through-the-web access to a large multivariate data archive, researchers are freed from the burdens of data storage and management. Additionally, the archive can simultaneously serve multiple users at multiple institutions in widely separated locations. Community-driven development of analysis routines allows users to visualize the data using both existing and custom-created code. E-science also encourages higher levels of collaboration, allowing researchers to form virtual teams able to tackle complex problems, where expertise in a variety of disciplines is required. Finally, by opening new avenues for interaction between researchers and students or members of the general public, e-science can influence both the questions scientists choose to address and the scope of their investigations. Transforming the promise of e-science into reality, however, is fraught with both technical and organizational challenges. The sheer volume of data records (50+ Tb/year) and observation density pose significant challenges for observatory and researcher alike, requiring new data mining approaches to be developed. Evolving and sometimes competing data format standards must be grappled with. Questions of data reliability and security must be answered. New protocols for protecting intellectual property within an open data environment must be defined. Ground rules for providing equitable access to finite shared resources (eg. underwater camera control time) must be defined. Cultural, institutional and motivational barriers to distributed decision-making and virtual team coordination must be overcome. NEPTUNE Canada is working to address the many challenges of e-science through a wide range of possible solutions. To help researchers make more optimal use of our large and growing data archives, we are developing a facility that allows users to upload and run custom data analysis routines on NEPTUNE Canada servers. Code authors will be able to retain privacy of over their routines, or if desired, publish their code for sharing and possible additional development with the larger user community. NEPTUNE Canada is developing other tools in the form of web and mobile applications for data search and subscription, event detection, interactive data plotting and real-time collaborative multi-user device control. Other custom tools will give users the ability to search, browse and annotate streaming media, then integrate and compile playlists from multiple sources to produce custom movies. Finally, the "glue" for an effective e-science working environment is under development in the form of web-based facilities to support and encourage project team coordination, communications, collaboration and electronic publication.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame distilled prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.

metaresearch head score (Codex)0.003
metaresearch head score (Gemma)0.000
Version: codex-gemma-dda1882f352aValidation status: machine_predicted_unvalidated
Candidate categoriesScholarly communication
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Observational · Consensus signal: none
GenreCandidate signal: Empirical · Consensus signal: Empirical
Teacher disagreement score0.522
Threshold uncertainty score1.000

Codex and Gemma teacher scores by category

CategoryCodexGemma
Metaresearch0.0030.000
Meta-epidemiology (narrow)0.0000.000
Meta-epidemiology (broad)0.0000.000
Bibliometrics0.0000.000
Science and technology studies0.0010.000
Scholarly communication0.0010.000
Open science0.0000.001
Research integrity0.0000.000
Insufficient payload (model declined to judge)0.0000.000

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.080
GPT teacher head0.365
Teacher spread0.285 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one teacher head, not a consensus.

Study designObservational
Domainnot available
GenreEmpirical

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations2
Published2010
Admission routes2
Has abstractyes

Explore more

Same topicScientific Computing and Data ManagementFrench-language works237,207