Bibliographic record
Abstract
Matthews, Joseph R. The Evaluation and Measurement of Library Services. Westport, CT: Libraries Unlimited, 2007. 372 pp. 50.00 USD. ISBN-10: 1- 59158-532-5. ISBN-13: 978-1-59158-532-9. 8 In this very substantial treatment of the question of how to measure and evaluate library services, Joseph Matthews has chosen to place the emphasis very much on evaluation. He laments the failure of many library administrators and directors to engage in meaningful evaluation of library services, their tendency to regard the gathering of statistical information as equivalent to evaluation, and their tendency to rely on the implicit goodness of libraries as justifications for the services they offer. The opening chapters of this book deal with evaluation issues and models as well as with the issues that arise from qualitative and quantitative forms of measurement and evaluation. In the subsequent chapters Matthews applies a variety of evaluation techniques to topics such as library users, the library collection, electronic resources, reference services, technical services, interlibrary loan, online systems, library instruction and information literacy, and customer service. The concluding chapters draw the reader from the specific to the more general: the economic and social impacts of libraries, communicating the value of library services to a wider audience, and methods to determine whether libraries provide life-long benefits to library users. Most chapters begin with a service definition, followed by a detailed discussion of the topic, a summary of the discussion, and very substantial footnotes and bibliographical information. One of the values of this book is the very wide range, geographically and historically, of evaluation and measurement studies that Matthews has consulted and his brief reflections on these studies. He notes areas where little research or evaluation has taken place and where more needs to be undertaken. Interestingly, Matthews does not include a chapter on library space as a service, and he misses the opportunity for an in-depth discussion of issues such as the use of library space for cultural events (poetry readings, book launchings, displays of artwork or handicrafts, etc.) or for human conveniences such as refreshment services, places for group study, or access to wireless Internet connections. There is a detailed discussion of electronic journals and e-books but curiously no discussion of other library electronic services such as Web site guides, pathfinders, or even the library's own Web site. The book is impressive in its breadth of coverage of this topic, in particular its discussion of the variety of types of evaluation and measurement. Whereas administrators of academic or public libraries will find much to benefit from in this book, administrators of special libraries may be disappointed at the limited discussion of the evaluation of the services they provide. …
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Direct model labels (unvalidated)
Per-model category and study-design labels from the labeling rounds. They are machine output, unvalidated, and the disagreement between models ships as data. No study design here is MEDLINE-validated yet.
| Model arm | Categories | Study design | Confidence |
|---|---|---|---|
| gemma | no category Domain: not available · Genre: Empirical About the Canadian research system: no · About a Canadian topic: no | Not applicable | low |
| gpt | no category Domain: not available · Genre: Empirical About the Canadian research system: no · About a Canadian topic: no | Observational | low |
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.047 | 0.119 |
| Meta-epidemiology (narrow) | 0.001 | 0.001 |
| Meta-epidemiology (broad) | 0.001 | 0.001 |
| Bibliometrics | 0.011 | 0.014 |
| Science and technology studies | 0.003 | 0.013 |
| Scholarly communication | 0.012 | 0.015 |
| Open science | 0.002 | 0.007 |
| Research integrity | 0.002 | 0.003 |
| Insufficient payload (model declined to judge) | 0.007 | 0.003 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedLabeled directly by 2 models reading the full record.
The models disagree on parts of this classification; every voice is preserved in the section at the end of the page.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".