Recent Developments in WTO Jurisprudence: Has the Appellate Body Resolved the Issue of an Appropriate Standard of Review in SPS Cases?
Bibliographic record
Abstract
A critical issue in World Trade Organization (WTO) dispute settlement is what standard of review dispute panels ought to apply to questions of fact when assessing the consistency of a country’s measures with trade rules. This question has arisen most acutely in the context of the SPS Agreement which requires countries to ensure that any SPS measure not based on international standards is not maintained without scientific evidence and is based on a risk assessment. This article explores how WTO dispute panels and the Appellate Body can strike a balance between a standard of review that, on the one hand, affords countries a degree of deference or flexibility and, on the other, ensures that they do not have open rein to justify any trade-restrictive measure on health-related grounds. Looking at recent cases, it finds that the Appellate Body has moved towards a balanced approach that requires panels to undertake a primarily procedural review and scrutinize the method by which the member reaches the decision in question, rather than focusing on the outcome of the decision.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.002 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.001 |
| Open science | 0.001 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.001 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".