Substandard and falsified medicine recalls in the legitimate supply chain: a systematic review of evidence
Bibliographic record
Abstract
OBJECTIVES: To compare international substandard and falsified (SF) medicine recall trends from published research papers based on governmental databases, summarise the extent of the problem in the legitimate supply chain, and identify ways to manage the issue. DESIGN: Systematic review of published academic evidence. DATA SOURCES: Drug recall data in published literature, obtained from official international government regulator databases in the USA, the UK, Canada, Sri Lanka, Zambia, Portugal, Nepal, Saudi Arabia, Argentina, Brazil, Chile, Cuba, Colombia, Mexico, Bolivia, Costa Rica, Ecuador, El Salvador, Guatemala, Honduras, Panama, Peru and Venezuela. ELIGIBILITY CRITERIA: A search for literature published between 2010 and 2024 was conducted using PubMed, MEDLINE and Embase. Included studies examined recall data of substandard and/or falsified medicines obtained through official government regulator websites. DATA EXTRACTION AND SYNTHESIS: Data were extracted using Excel files and synthesised using a thematic analysis approach. RESULTS: 13 research papers containing original data were included. Recall data were obtained from official regulatory databases in 23 different countries. Substandard medicines had significantly higher recall rates than falsified products, while parenteral drugs and tablets were the most recalled formulation types. The leading reasons for defective medicines were contamination, out-of-specification results, stability and packaging issues. India was identified as a common source of SF medicines in Zambia, Sri Lanka, Brazil and Nepal. Frequent recalls of anti-infective drugs were observed in countries with equatorial, tropical and subtropical climates, while high-income countries like Canada, Saudi Arabia and the UK faced issues with defective antihypertensive drugs. Interestingly, medicines affected by nitrosamines' contamination were recalled in all regions examined in 2018, but in other recall cases, there were disparities among recall action. CONCLUSIONS: There appeared to be similar international recall practices for some products like nitrosamines and not for others like rosiglitazone across the same time frame, which raises questions concerning international drug safety disparities. The requirement to align and globally strengthen regulatory frameworks was of emerging importance. Cooperation between regulatory authorities to create a harmonised approach to reporting medicine recalls and standardising the data included in a recall notification is proposed to facilitate a more accurate comparison of international trends surrounding recalled SF medicines.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.037 | 0.142 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.009 | 0.008 |
| Bibliometrics | 0.025 | 0.023 |
| Science and technology studies | 0.001 | 0.002 |
| Scholarly communication | 0.005 | 0.006 |
| Open science | 0.003 | 0.003 |
| Research integrity | 0.003 | 0.002 |
| Insufficient payload (model declined to judge) | 0.004 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".