Bibliographic record
Abstract
In Reply: I love these two letters! They are another important point and counterpoint for this important discussion. Let me be the first to concede and agree with Gliatto et al that my suggestion to have someone outside of the dean’s office prepare the Medical Student Performance Evaluation (MSPE) is not realistic; however, it did get their attention. They also make the observation that it is not who writes the letter but how the letter is written, emphasizing that objectivity follows if schools use the guidelines published in 1989.1 Weissman makes a strong argument, even if overstated, that if these letters don’t improve in objectivity, they will continue to be ignored. Regarding how the letters are written, the MSPE has shown slow but steady improvement in the number of schools that adhere to these important guidelines. 2–4 The percentage of MSPE letters assessed as adequate, using the guidelines as the benchmark, has increased from 55% (1993), to 65% (1998), and to 75% in the most recent analysis (2005). The authors of the latest article noted a particular weakness across many schools in that only 17% provided “stand alone” comparative information in the summary paragraph; the authors commented, “Herein lies the Achilles heel of the MSPE; the dean of students must advocate, and the program directors must discriminate.” So, allow me to make a more realistic recommendation about who writes the letters and an outlandish recommendation about how the letters are written. For the who, let’s follow the example of our Liaison Committee on Medical Education-accredited medical schools in Canada, where every one of the 17 schools relies on the education office’s side of the dean’s office to create the MSPE. That is, after all, the evaluation side of the dean’s office, as opposed to the advocates on the student affairs side. The outlandish recommendation about the how is to ask all residency directors to give notice that in two years, they will no longer offer interviews to students who come from medical schools that refuse to follow the national guidelines. This would be totally unfair to students and therefore beyond outlandish and into unacceptable. However, regardless of who writes the letters, we cannot tolerate that a sizable percentage of schools bring the other schools’ good efforts down. Dan Hunt, MD, MBA Co-secretary, Liaison Committee on Medical Education, and senior director, Accreditation Services, Association of American Medical Colleges, Washington, D.C.; [email protected]
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.001 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".