MétaCan
Menu
Retour à la cohorte
Enregistrement W2980077923 · doi:10.1353/vpr.2019.0037

Indexing, Checking, and Encoding in the Periodical Poetry Index

2019· article· en· W2980077923 sur OpenAlexvenueaboutno aff
Natalie M. Houston, Lindsy Lawrence, April Patrick

Notice bibliographique

RevueVictorian periodicals review · 2019
Typearticle
Langueen
DomaineArts and Humanities
ThématiqueDigital Humanities and Scholarship
Établissements canadiensnon disponible
Organismes subventionnairesnon disponible
Mots-clésIndex (typography)ScholarshipPoetryPeriodical literatureSearch engine indexingHistoryLibrary scienceSociologyLiteratureComputer scienceArtInformation retrievalLawWorld Wide WebPolitical science

Résumé

récupéré en direct d'OpenAlex

Indexing, Checking, and Encoding in the Periodical Poetry Index Natalie M. Houston (bio), Lindsy Lawrence (bio), and April Patrick (bio) This cluster of essays originates from a panel at the Research Society for Victorian Periodicals and Victorian Studies Association of Western Canada joint conference in July 2018. On this panel we explored some of the different scholarly activities that compose the Periodical Poetry Index, discussing the theoretical underpinnings and methodological commitments of our work that would interest both scholars in periodical studies and the digital humanities more broadly. The breadth of bibliographic information now presented by the Periodical Poetry Index comes from our discoveries made while indexing, checking, and encoding the bibliographic data from Blackwood's Edinburgh Magazine (1817–1900), the Cornhill (1860–1900), Dark Blue (1871–73), and Macmillan's Monthly Magazine (1859–1900).1 There is a long tradition of collaborative bibliographic scholarship in periodical studies. Most notably The Wellesley Index to Victorian Periodicals, published in five volumes from 1966 to 1988, has served as the foundation for modern research in Victorian periodicals. The idea for The Wellesley Index originated in correspondence between Walter Houghton and Richard Altick as they shared their knowledge of periodical contributors writing unsigned articles.2 Houghton saw the tremendous resources available to scholars in Victorian periodicals and sought to make them more accessible by creating a research tool that reproduced the tables of contents and identified the authors of a large number of unsigned pieces. Like the Wellesley team, we benefit from scholarly collaboration that spans geographic location, and we have the added benefits of digital tools that support our communication and research. Sharing our work digitally frees us from the physical and financial constraints that often burdened Houghton and his team, who had to persuade the University of Toronto [End Page 604] Press to continue publishing an ever-expanding scholarly resource. Houghton recalled, "In my correspondence with the University of Toronto Press, I had agreed to two volumes and not over 100,000 items in Part A—fewer if possible," but his initial agreements with the press would be greatly surpassed as the project continued.3 The scale of Victorian periodical publishing was much greater than even Houghton had initially imagined. As work on The Wellesley Index proceeded, the editors not only covered additional periodical titles but also incorporated updates and corrections in later volumes. Eileen Curran's updates to the Wellesley information, published in regular intervals in Victorian Periodicals Review and later as The Curran Index, continued this iterative process of discovery and correction. Because we are working with digital surrogates created by the Google Books project, we can conduct most of our research without traveling to libraries with large periodical holdings. As we make corrections or alterations to the data collected for our project, we can update the database and the web display of its information. Our project was inspired by Linda K. Hughes's essay "What the Wellesley Index Left Out: Why Poetry Matters to Periodical Studies," which vividly demonstrates the scholarly lacunae created by Houghton's decision to omit poetry from the indexes produced by his team. Because The Wellesley Index made Victorian periodicals more accessible to scholars, poetry's "omission skews our understanding of both Victorian poetry and periodicals, which were interrelated in highly complex terms."4 Although aesthetic considerations undoubtedly played a part in this decision, so did the Wellesley's focus on identifying unsigned contributors: Houghton suggests that "to include 7000 or so poems, in many cases anonymous or pseudonymous, and if signed, by obscure versifiers, would cost far too much space, labor, and funding."5 He greatly underestimated the sheer quantity of poems published in the pages of Victorian periodicals. Our research in five titles alone has uncovered almost 5,000 poems, and we fully expect to index thousands more in the years to come. In the first phase of our project, we are indexing poems published in the forty-five periodicals covered in The Wellesley Index. We know that approximately half of these titles contain significant quantities of verse, but the exact details of poetry's distribution will only be evident once we have collected data for each of the titles. Because The...

Récupéré en direct depuis OpenAlex et désinversé. Les résumés ne sont pas conservés dans cette base de données : les index inversés représentent 8,6 Go des 9,3 Go de texte de la base, et le serveur dispose de 13 Go libres.

Comment cette classification a été obtenuedéplier

Prédiction machine sur la base complète

Imitation des enseignants

Ni prévalence calibrée, ni vérité terrain. Validation humaine à venir. Le volet Gemma est une étiquette directe du modèle pour chaque travail de la base, lue sur la notice réduite au titre. Le volet Codex est un classifieur appris des 10 348 étiquettes directes de Codex et calibré sur les taux pondérés de l'échantillon; les champs sans appui suffisant ne portent aucun appel Codex. Le mode candidate est l'union des deux volets; le consensus est leur intersection. Ces sorties portent le statut machine_predicted_unvalidated et ne sont pas des étiquettes humaines.

score de la tête « metaresearch » (Codex)0,017
score de la tête « metaresearch » (Gemma)0,144
Version: metacan-v3-hybrid-931329e0061cStatut de validation: machine_predicted_unvalidated
Catégories candidatesaucune
Catégories consensuellesaucune
DomaineSignal candidat: aucune · Signal consensuel: aucune
Devis d'étudeSignal candidat: Sans objet · Signal consensuel: aucune
GenreSignal candidat: Empirique · Signal consensuel: aucune
Score de désaccord entre enseignants0,022
Score d'incertitude au seuil0,092

Scores du classifieur distillé par catégorie (deux têtes)

CatégorieCodexGemma
Métarecherche0,0170,144
Méta-épidémiologie (sens strict)0,0000,001
Méta-épidémiologie (sens large)0,0010,000
Bibliométrie0,0150,022
Études des sciences et des technologies0,0080,019
Communication savante0,0220,025
Science ouverte0,0020,009
Intégrité de la recherche0,0010,005
Charge utile insuffisante (le modèle a refusé de juger)0,0150,006

Scores machine (provisoires)

Les deux têtes enseignantes du modèle étudiant, lues sur ce travail. Un score ordonne la base pour la relecture; il n'affirme jamais une catégorie, et le statut de validation accompagne chaque rangée tel quel.

Scores de référence d'un modèle non mature (critères de maturité non atteints, 7 itérations). Un score ordonne; il n'affirme jamais une catégorie.

Tête enseignante Opus0,029
Tête enseignante GPT0,255
Écart entre enseignants0,226 · la distance entre les deux têtes enseignantes sur ce seul travail
Statut de validationscore_only:v0-immature-baseline · tel quel depuis la passe de notation : score_only signifie que le nombre peut ordonner les travaux, et qu'aucune étiquette de catégorie n'en découle

Classification

machine, non validée

Prédiction automatique; un appel candidat d’une seule source (Gemma direct ou Codex distillé), pas un consensus.

Les modèles n’ont appliqué aucune catégorie : rien dans la taxonomie ne correspondait à ce travail.
Devis d'étudeSans objet
Domainenon disponible
GenreEmpirique

Le détail, modèle par modèle et score par score, se trouve en fin de page sous « Comment cette classification a été obtenue ».

En bref

Citations0
Publié2019
Routes d'admission2
Résumé présentoui

Explorer davantage

Même revueVictorian periodicals reviewMême sujetDigital Humanities and ScholarshipTravaux en français237 207