Curating Archaeological Knowledgein the Digital Continuum: from Practiceto Infrastructure
Bibliographic record
Abstract
Abstract As a “grand challenge” for digital archaeology, I propose the adoption of programmatic research to meet the challenges of archaeological curation in the digital continuum, contingent on curation-enabled global digital infrastructures, and on contested regimes of archaeological knowledge production and meaning making. My motivation stems from an interest in the sociotechnical practices of archaeology, viewed as purposeful activities centred on material traces of past human presence. This is exemplified in contemporary practices of interpretation “at the trowel’s edge”, in epistemological reflexivity and in pluralization of archaeological knowledge. Adopting a practice-centred approach, I examine how the archaeological record is constructed and curated through archaeological activity “from the field to the screen” in a variety of archaeological situations. I call attention to Çatalhöyük as a salient case study illustrating the ubiquity of digital curation practices in experimental, well-resourced and purposefully theorized archaeological fieldwork, and I propose a conceptualization of digital curation as a pervasive, epistemic-pragmatic activity extending across the lifecycle of archaeological work. To address these challenges, I introduce a medium-term research agenda that speaks both to epistemic questions of theory in archaeology and information science, and to pragmatic concerns of digital curation, its methods, and application in archaeology. The agenda I propose calls for multidisciplinary, multi-team, multiyear research of a programmatic nature, aiming to re-examine archaeological ontology, to conduct focused research on pervasive archaeological research practices and methods, and to design and develop curation functionalities coupled with existing pervasive digital infrastructures used by archaeologists. It has a potential value in helping to establish an epistemologically coherent framework for the interdisciplinary field of archaeological curation, in aligning archaeological ontologies work with practice-based, agencyoriented and participatory theorizations of material culture, and in matching the specification and design of archaeological digital infrastructures with the increasingly globalized, ubiquitous and pervasive digital information environment and the multiple contexts of contemporary meaning-making in archaeology.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.030 | 0.026 |
| Meta-epidemiology (narrow) | 0.000 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.003 | 0.003 |
| Science and technology studies | 0.009 | 0.071 |
| Scholarly communication | 0.020 | 0.023 |
| Open science | 0.004 | 0.027 |
| Research integrity | 0.004 | 0.005 |
| Insufficient payload (model declined to judge) | 0.007 | 0.001 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".