Re: "Examination of How Neighborhood Definition Influences Measurements of Youths' Access to Tobacco Retailers: A Methodological Note on Spatial Misclassification"
Bibliographic record
Abstract
In a recent article, Duncan et al. (1) compared tobacco retailer density and proximity measures computed within 2 administrative and 4 egocentric neighborhood definitions and precisely identified which measures significantly differed across neighborhood definitions. This comparative analysis led the authors to conclude, rightly so, that how one defines “neighborhood” may considerably influence neighborhood-level exposure measures. These findings echo the importance of both zoning and spatial scale, which has often been underlined in the health literature (2–4) since Openshaw first defined the Modifiable Areal Unit Problem in 1984 (5). The authors went on to conclude that “whenever possible, egocentric neighborhood definitions should be used” and that “the use of larger administrative neighborhood definitions can bias exposure estimates for proximity” (1). However, the authors should not have extended their findings, which concerned the difference between neighborhood resource accessibility measures, to a judgment on the most relevant spatial units to use in neighborhood and health research. We believe their conclusion stems from the following 2 assumptions that are pervasive in the literature and deserve to be discussed: that “egocentric is better” and that “smaller is better.” First, when concluding that egocentric neighborhoods should be preferred, the authors assume that isotropic areas (i.e., spreading out uniformly in all directions around individuals' homes) are necessarily the optimal way to delineate neighborhoods (6). However, historical, social, and political processes may prevent people from experiencing certain places and reaching specific resources despite being their located close to their homes. Administrative areas, which are, by definition, not centered on individuals' homes, may in some cases provide more adequate estimates of neighborhood resource accessibility than egocentric areas would, notably when they have been delineated by taking historical, social, and political divisions into account. There is a real need to discuss the unjustified use of administrative areas to define neighborhoods. However, one should not fall into the opposite extreme by claiming that egocentric neighborhoods are necessarily better. Second, the authors assume that using administrative areas larger than census tracts would inevitably increase the likelihood of spatial misclassification. Actually, depending on the profile and location of individuals in a city, it could be more relevant to derive neighborhood exposure measures from units larger than census tracts. In a study in the Paris, France, metropolitan area, people's health-seeking behaviors were better modeled by neighborhood resource densities computed from groups of adjacent census tracts than from the residential census tract only (4). Areas larger than census tracts have also been found to better fit with inner-Paris inhabitants' perceived neighborhoods (7). We therefore urge neighborhood and health researchers to justify their choice of a given neighborhood definition by comparing it with validity criteria involving people. Although there is no “gold standard,” the following 3 people-based criteria may be distinguished:The above are mere suggestions, but whichever validity criterion is chosen, we strongly recommend that researchers refer to people, and not only to places, before concluding that there is spatial misclassification in neighborhood exposures. The most common approach is to undertake sensitivity analyses by correlating neighborhood exposure measures with people's health indicators to identify the neighborhood definition that maximizes measures of association (4, 8, 9). Investigating the correlation between people's subjective assessments of neighborhood resources and objective measures of these same resources in various spatial units (10) is also a promising avenue for selecting spatial units that closely approximate people's assessments (6). It may be relevant to choose a neighborhood delineation that fits people's neighborhood experiences. Cognitive mapping, which has highlighted that perceived neighborhoods may vary greatly in spatial extent according to people's profile and location, may, for instance, provide invaluable insights for comparing neighborhood definitions (7, 11). The authors are supported by the Département de Médecine Sociale et Preventive of the Université de Montréal, the Institut de Recherche en Santé Publique of the Université de Montréal, and the Spatial Health Research Lab in CHUM Research Centre. J.V. was also supported by the Centre National de la Recherche Scientifique in France. Conflict of interest: none declared.
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame machine prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. The Gemma side is a direct model label for every work in the frame, read from the title-only record. The Codex side is a classifier learned from the 10,348 direct Codex labels and calibrated to design-weighted sample rates; fields without enough sample support carry no Codex call. Candidate is the union of the two sides; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels.
Distilled classifier scores by category (both heads)
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.035 | 0.179 |
| Meta-epidemiology (narrow) | 0.002 | 0.002 |
| Meta-epidemiology (broad) | 0.003 | 0.003 |
| Bibliometrics | 0.002 | 0.002 |
| Science and technology studies | 0.011 | 0.009 |
| Scholarly communication | 0.007 | 0.006 |
| Open science | 0.008 | 0.005 |
| Research integrity | 0.098 | 0.092 |
| Insufficient payload (model declined to judge) | 0.006 | 0.007 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one source (direct Gemma or distilled Codex), not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".