MétaCan
Menu
Back to cohort
Record W2074279928 · doi:10.4000/jtei.210

‘The Apex of Hipster XML GeekDOM’

2011· article· en· W2074279928 on OpenAlexaff
Lynne Siemens, Ray Siemens, Hefeng Wen, Cara Leitch, Dot Porter, Liam Sherriff, Karin Armstrong, Melanie Chernyk

Bibliographic record

VenueJournal of the Text Encoding Initiative · 2011
Typearticle
Languageen
FieldArts and Humanities
TopicDigital Humanities and Scholarship
Canadian institutionsUniversity of Victoria
Fundersnot available
KeywordsDigital humanitiesField (mathematics)SociologyCommonsEpistemologyVirtueDigital libraryComputer scienceLibrary sciencePolitical scienceLinguisticsPhilosophyLaw

Abstract

fetched live from OpenAlex

If the notion of the methodological commons is as centrally located as we believe it to be in any visualization accurately depicting the intellectual structure of the digital humanities and digital literary studies (McCarty 2005, 119), then so, too, must be the community itself whose members provide that which populates the commons. As an interdiscipline, humanities computing has always well-understood its methodologies; indeed, the digital humanities (of which digital literary studies is a part), more generally, have made a virtue of the way in which they render explicit and tangible the theoretical models that govern the representative and analytical endeavour of their fields via computational application. So, too, have those in the field understood and documented its formal structures and institutional manifestations, a chief example being the Text Encoding Initiative itself. Less explicitly rendered and less formally documented–though intuited by its chief practitioners and builders–is the exact nature of the community itself, its depth and breadth, its own centre and, perhaps more important in a field whose embrace of interdisciplinarity is far from self-serving, its periphery and those aspects of which promise to become central. This article presents work carried out in conjunction with the Text Encoding Initiative Consortium, a foundation of many digital literary studies projects, work that seeks to document the full nature of its community, from the institutional and research project groups that comprise the formal consortium at centre to those who appear on the other side of the easily-permeable periphery that separates it from the centre, largely individual practitioners in areas hitherto not closely identified with the digital humanities but clearly sharing methods and tools, thus suggesting their place in the same communities of practice, as they are members of the same methodological commons. This methodological approach is drawn from marketing and organizational behavior, manifest in social networking, in the study of viral marketing campaigns conducted in online environments. The method for this work was centred around a viral marketing experiment designed to showcase the TEI and novel ways that it can be used to encode different kinds of text. At the heart of the experiment was a Bob Dylan song and its associated video which incorporated text; encoded text was overlaid and the video was posted to YouTube and a blog with links to the TEI website with analysis of traffic patterns carried out.

Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.

How this classification was reachedexpand

Full frame machine prediction

Teacher imitation

Not calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.

metaresearch head score (Codex)0.008
metaresearch head score (Gemma)0.030
Version: metacan-v3-hybrid-931329e0061cValidation status: machine_predicted_unvalidated
Candidate categoriesnone
Consensus categoriesnone
DomainCandidate signal: none · Consensus signal: none
Study designCandidate signal: Not applicable · Consensus signal: none
GenreCandidate signal: Other · Consensus signal: none
Teacher disagreement score0.062
Threshold uncertainty score0.208

Distilled classifier scores by category (both heads)

CategoryCodexGemma
Metaresearch0.0080.030
Meta-epidemiology (narrow)0.0010.001
Meta-epidemiology (broad)0.0010.001
Bibliometrics0.0030.003
Science and technology studies0.0020.004
Scholarly communication0.0130.023
Open science0.0040.011
Research integrity0.0030.005
Insufficient payload (model declined to judge)0.0620.032

Machine scores (provisional)

The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.

Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.

Opus teacher head0.155
GPT teacher head0.246
Teacher spread0.091 · how far apart the two teachers sit on this one work
Validation statusscore_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from it

Classification

machine, unvalidated

Machine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.

The models applied no category: nothing in the taxonomy fit this work.
Study designNot applicable
Domainnot available
GenreOther

How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".

Quick stats

Citations1
Published2011
Admission routes1
Has abstractyes

Explore more

Same venueJournal of the Text Encoding InitiativeSame topicDigital Humanities and ScholarshipFrench-language works237,207