Skip to main page content
U.S. flag

An official website of the United States government

Dot gov

The .gov means it’s official.
Federal government websites often end in .gov or .mil. Before sharing sensitive information, make sure you’re on a federal government site.

Https

The site is secure.
The https:// ensures that you are connecting to the official website and that any information you provide is encrypted and transmitted securely.

Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2020 Jan 28;20(1):15.
doi: 10.1186/s12874-019-0894-6.

Sample size issues in time series regressions of counts on environmental exposures

Affiliations

Sample size issues in time series regressions of counts on environmental exposures

Ben G Armstrong et al. BMC Med Res Methodol. .

Abstract

Background: Regression analyses of time series of disease counts on environmental determinants are a prominent component of environmental epidemiology. For planning such studies, it can be useful to predict the precision of estimated coefficients and power to detect associations of given magnitude. Existing generic approaches for this have been found somewhat complex to apply and do not easily extend to multiple series studies analysed in two stages. We have sought a simpler approximate approach which can easily extend to multiple series and give insight into factors determining precision.

Methods: We derive approximate expressions for precision and hence power in single and multiple time series studies of counts from basic statistical theory, compare the precision predicted by these with that estimated by analysis in real data from 51 cities of varying size, and illustrate the use of these estimators in a realistic planning scenario.

Results: In single series studies with Poisson outcome distribution, precision and power depend only on the usable variation of exposure (i.e. that conditional on covariates) and the total number of disease events, regardless of how many days those are spread over. In multiple time series (eg multi-city) studies focusing on the meta-analytic mean coefficient, the usable exposure variation and the total number of events (in all series) are again the sole determinants if there is no between-series heterogeneity or within-series overdispersion. With heterogeneity, its extent and the number of series becomes important. For all but the crudest approximation the estimates of standard errors were on average within + 20% of those estimated in full analysis of actual data.

Conclusions: Predicting precision in coefficients from a planned time series study is possible simply and given limited information. The total number of disease events and usable exposure variation are the dominant factors when overdispersion and between-series heterogeneity are low.

Keywords: Environment; Poisson regression; Power; Sample size; Statistics; Time series regression.

PubMed Disclaimer

Conflict of interest statement

The authors declare that they have no competing interests.

Figures

Fig. 1
Fig. 1
Approximate power as a function of number of cases, size of risk to be detected in relation to usable exposure spread, and overdispersion. Blue lines (left cluster): coefficient (log(RR) per useable SD(x|z)) = 5%. Green lines (middle-left cluster): coefficient = 2% per SD(x|z). Red lines (middle right cluster): coefficient = 1% per SD(x|z). Black lines (right cluster): coefficient = 0.05% per SD(x|z). Within clusters, from left (top): Solid (left): dispersion = 1.0. Dashed (middle): dispersion = 1.2. Dotted (right): dispersion = 1.5

References

    1. Atkinson R, Kang S, Anderson H, Mills I, Walton H. Epidemiological time series studies of PM2. 5 and daily mortality and hospital admissions: a systematic review and meta-analysis. Thorax. 2014;69(7):660–665. doi: 10.1136/thoraxjnl-2013-204492. - DOI - PMC - PubMed
    1. Atkinson RW, Mills IC, Walton HA, Anderson HR. Fine particle components and health—a systematic review and meta-analysis of epidemiological time series studies of daily mortality and hospital admissions. J Expo Science Environ Epidemiol. 2015;25(2):208–214. doi: 10.1038/jes.2014.63. - DOI - PMC - PubMed
    1. Bhaskaran K, Gasparrini A, Hajat S, Smeeth L, Armstrong B. Time series regression studies in environmental epidemiology. Int J Epidemiol. 2013;42:1187–1195. doi: 10.1093/ije/dyt092. - DOI - PMC - PubMed
    1. Ye X, Wolff R, Yu W, Vaneckova P, Pan X, Tong S. Ambient temperature and morbidity: a review of epidemiological evidence. Environ Health Perspect. 2012;120(1):19–28. doi: 10.1289/ehp.1003198. - DOI - PMC - PubMed
    1. Yu W, Mengersen K, Wang X, Ye X, Guo Y, Pan X, Tong S. Daily average temperature and mortality among the elderly: a meta-analysis and systematic review of epidemiological evidence. Int J Biometeorol. 2012;56(4):569–581. doi: 10.1007/s00484-011-0497-3. - DOI - PubMed