The hard art of soft science: Evidence‐Based Medicine, Reasoned Medicine or both?
Bibliographic record
Abstract
In the past 14 years, Evidence-Based Medicine (EBM) has enjoyed unprecedented developments and gained widespread acceptance among health professionals. However, should we be content with producing, critically appraising and using the best evidence available for our understanding of health problems and decision making about them? Are our convictions about EBM's relevance, our conviction and intellectual satisfaction with its mastery and adoption enough? Should we continue pushing forward along this promising path, or should we further diversify the content and scope of EBM? Is EBM the only way to view medicine in the near future? This paper presents some options to choose from in terms of direction and content as well as questions to answer given the current EBM crossroads. More intensive and extensive EBM combined with 'other features'-based medicines may be the preferred strategy to follow in the future to determine the development, use and evaluation of EBM. Argument-based medicine or Reasoned Medicine is one of the options that can be integrated into the mainstream of medical reasoning and decision making.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.372 | 0.792 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.002 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.000 | 0.003 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.003 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; both teacher heads agree on what is shown here.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".