Analysis of British American Tobacco's questionable use of privilege and protected document claims at the Guildford Depository
Bibliographic record
Abstract
BACKGROUND: Tobacco companies have a documented history of attempting to hide information from public scrutiny, including inappropriate privilege claims. The 1998 Minnesota Consent Judgement created two depositories to provide public access to discovered documents. Users raised concerns about the access conditions and ongoing integrity of the Guildford Depository collection operated until 2015 by British American Tobacco (BAT). METHODS: A metadata search of the Legacy Tobacco Documents Library identified inconsistent privilege claims, and duplicates of documents withheld by BAT from public visitors. A review of the validity of claims, for documents obtained through these searches, was conducted against recognised legal definitions of privilege. FINDINGS: BAT has asserted inappropriate privilege claims over 49% of the documents reviewed (n=63). The quantity of such claims and consistency of the stated rationale for the privilege claims suggest a concerted effort rather than human error. CONCLUSIONS: There was insufficient attention given to the operation of the Guildford Depository by the original plaintiffs, including to the subsequent use of privilege claims. Appropriate access to these documents, commensurate with the terms of legal settlements creating the collection, was critical given their public interest value for enhancing understanding of industry strategies and activities, informing of policy interventions, and for holding the industry to account. Future legal settlements should prevent defendants from subsequently withholding disclosed documents, aside from those legitimately privileged, from public view. Control of publicly disclosed documents should not be placed back into the hands of defendant tobacco companies. Plaintiffs also need to invest adequate resources into policing claims of legal privilege.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.000 | 0.000 |
| Bibliometrics | 0.000 | 0.000 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.000 |
| Insufficient payload (model declined to judge) | 0.000 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".