Bibliographic record
Abstract
Ostraka in the Collection of New York University is a comprehensive edition and commentary of 77 ostraka, or potsherds with ancient texts written on them, from Greco-Roman and late antique Egypt. Seventy-two of these ostraca are housed in NYU Special Collections, originally purchased by Caspar Kraemer in 1932, then the chair of the NYU Classics Department. Although Kraemer advertised the imminent publication of the texts in 1934 and later collaborated with the famed papyrologist Herbert Youtie, neither completed the project. The ostraka in this small collection span the 2nd century BCE to the 8th century CE and include both Greek and Coptic texts. The majority, however, form a coherent dossier of tax receipts related to mortuary activities in Upper Egypt during the reign of Augustus (texts 7-70, dated from roughly the last quarter of the 1st century BCE to 12 CE). The five ostraka published in this volume not held by NYU include one that had been part of Kraemer’s original purchase but was subsequently lost (thankfully preserved in a photograph in Youtie’s archive at the University of Michigan), and four ostraka now held by the Los Angeles County Museum of Art. The latter four texts were purchased separately and published previously, but clearly belong to the same group of texts. They are included in this volume both for the sake of completeness and because the present authors were able to improve the readings in light of the context provided by the dossier as a whole. In addition to the scholarly edition of these texts, the volume contains a full discussion of their provenance, the taxes involved, the taxpayers and tax-collectors, and a ceramological analysis of the sherds as media for these texts.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.001 | 0.005 |
| Meta-epidemiology (narrow) | 0.001 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.003 | 0.003 |
| Science and technology studies | 0.009 | 0.005 |
| Scholarly communication | 0.007 | 0.004 |
| Open science | 0.001 | 0.005 |
| Research integrity | 0.001 | 0.003 |
| Insufficient payload (model declined to judge) | 0.109 | 0.017 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".