What qualities are rare in examiners reports?
Bibliographic record
Abstract
For research students in Australia and other nations the PhD thesis is the pinnacle of higher degree endeavour. Unlike most other nations, however, the written report on the thesis is the only assessment that most Australian candidates are likely to receive. In the USA coursework provides a substantial component of assessment, while in the UK the viva voce is required. Although doctoral coursework is gaining ground in Australia, thesis examination remains the dominant form of assessment. Moreover the likelihood of moving to a viva is slim, especially as serious concerns about the credibility of the oral examination are emerging. In Australia research student enrolments numbered 37,175 in 1999 and total completions for the previous year was 5,109. This means that between 10,000 and 15,000 examiners reports are required annually. Despite the importance, scope and intensity of the process the topic of thesis examination has rarely attracted research interest. However, in a climate of quality assurance and high research competitiveness this already changing, as evident in the increasing number of studies emerging from the UK. The examiners' written reports on research theses are idiosyncratic and individualistic documents, despite efforts to standardise or structure them. All manner of reasons can be advanced for the characteristics of the written report ranging from the unusual nature of the assessment task itself, through to the lack of funds devoted to its execution. However, of particular interest in regard to exploring the quality of both academic outcomes and examination process is what examiners regard as important enough to include in the report, how they communicate this information and what both the content and the sub-text reveals about their expectations. This paper concentrates on what topics and qualities of comment are unusual or relatively sparse in examiners' reports on PhD theses. The findings are based on the core content analysis of 303 examiners reports on 101 candidates at one NSW university with a strong research profile.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.016 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".