Back

popexposure: An open-source Python package to find the number of people residing near environmental hazards quickly and efficiently

McBrien, H.; Casey, J. A.; Chillrud, L. G.; Flores, N. M.; Wilner, L. B.

2025-10-22 public and global health
10.1101/2025.10.19.25338326 medRxiv
Show abstract

Environmental scientists often assess exposure to hazards using residential proximity (i.e., they consider an individual living near a hazard to be exposed). Such assessment requires large, fine-scale spatial datasets that describe locations of environmental hazards and residential populations. Manipulating such datasets is technically demanding, slow, memory-intensive, and difficult to optimize for speed and memory use. Currently, individual research teams each write their own algorithms for this task. This may lead to inconsistencies in assumptions, methods, and results. We developed an open-source Python package, popexposure, which quickly, efficiently, and consistently estimates the number of people living near environmental hazards. Given a set of distinct hazard geometries and corresponding buffer distances, popexposure can estimate the number of people living within the buffered area of each hazard using a gridded population dataset. popexposure can also estimate the number of people living within the buffer distance of each hazard by additional administrative geographies. For example, users can calculate the number of people exposed to hazards in each census tract or zip code tabulation area (ZCTA). popexposure addresses common issues encountered in this calculation: whether or not to double-count people exposed to more than one hazard, proper pixel apportionment, choosing appropriate map projections for data covering large areas, and optimizing speed and memory. In this paper, we describe popexposures functionality and provide an example use case, calculating the proportion of people exposed to any wildfire burn zone disaster in California in 2018 in each ZCTA. What this study addsEnvironmental epidemiologists often assess exposure to hazards using residential proximity (i.e., they consider an individual exposed if they live near a hazard). This computation presents technical difficulties, and different research teams each apply their own solution, since no software currently exists to do this task. We developed an open-source Python package, popexposure, which quickly, efficiently, and consistently estimates the number of people living near environmental hazards. Here, we describe the package and provide an example use case, applying popexposure to compute the proportion of people exposed to any wildfire burn zone disaster in California in 2018 in each ZCTA.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

1
PLOS ONE
5266 papers in training set
Top 8%
19.7%
2
PLOS Computational Biology
1863 papers in training set
Top 3%
11.7%
3
GeoHealth
12 papers in training set
Top 0.1%
4.6%
4
Environmental Health Perspectives
17 papers in training set
Top 0.1%
4.3%
5
BMC Medical Research Methodology
47 papers in training set
Top 0.3%
4.3%
6
Methods in Ecology and Evolution
176 papers in training set
Top 0.8%
2.6%
7
PLOS Global Public Health
344 papers in training set
Top 5%
2.6%
8
Scientific Reports
3612 papers in training set
Top 40%
2.6%
50% of probability mass above
9
PeerJ
308 papers in training set
Top 3%
2.5%
10
eLife
5828 papers in training set
Top 39%
2.5%
11
Environmental Research
49 papers in training set
Top 0.5%
2.3%
12
Open Research Europe
14 papers in training set
Top 0.1%
2.3%
13
American Journal of Epidemiology
67 papers in training set
Top 0.6%
1.8%
14
Patterns
78 papers in training set
Top 2%
1.2%
15
PLOS Water
13 papers in training set
Top 0.2%
1.2%
16
JMIR Public Health and Surveillance
45 papers in training set
Top 1%
1.2%
17
Ecological Informatics
33 papers in training set
Top 0.5%
1.1%
18
Bioinformatics
1204 papers in training set
Top 8%
1.1%
19
Epidemics
116 papers in training set
Top 2%
1.1%
20
Annals of Epidemiology
21 papers in training set
Top 0.5%
0.9%
21
Science of The Total Environment
186 papers in training set
Top 3%
0.9%
22
International Journal of Environmental Research and Public Health
128 papers in training set
Top 5%
0.9%
23
Infectious Disease Modelling
54 papers in training set
Top 1%
0.9%
24
Nature Communications
5641 papers in training set
Top 58%
0.6%
25
BMC Bioinformatics
457 papers in training set
Top 6%
0.6%
26
Data in Brief
14 papers in training set
Top 0.1%
0.6%
27
iScience
1154 papers in training set
Top 42%
0.5%
28
SoftwareX
15 papers in training set
Top 0.4%
0.5%
29
Biology Methods and Protocols
61 papers in training set
Top 3%
0.5%
30
Ecology and Evolution
267 papers in training set
Top 6%
0.5%