PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 13, 2007PLoS ONE244 citationsOpen Access

Maximum Likelihood Estimation of the Negative Binomial Dispersion Parameter for Highly Overdispersed Data, with Applications to Infectious Diseases

JLJames O. Lloyd‐Smith

Key Points

  • To examine the bias and precision of maximum-likelihood estimates of the negative binomial dispersion parameter, k, in highly overdispersed datasets.
  • Conducted a simulation study to explore bias, precision, and confidence interval coverage of estimates for k.
  • Analyzed datasets influenced by event under-counting and selection bias during disease outbreaks.
  • Assessed small-sample bias and estimation accuracy for various event reporting scenarios.
  • Maximum likelihood estimates of k show upward bias in small samples or with zero-class event under-reporting.
  • Confidence intervals based on asymptotic variance consistently fall below the nominal coverage level.
  • Estimation from outbreak datasets does not increase bias for k estimates, though it may inflate mean estimates.

Abstract

BACKGROUND: The negative binomial distribution is used commonly throughout biology as a model for overdispersed count data, with attention focused on the negative binomial dispersion parameter, k. A substantial literature exists on the estimation of k, but most attention has focused on datasets that are not highly overdispersed (i.e., those with k>or=1), and the accuracy of confidence intervals estimated for k is typically not explored. METHODOLOGY: This article presents a simulation study exploring the bias, precision, and confidence interval coverage of maximum-likelihood estimates of k from highly overdispersed distributions. In addition to exploring small-sample bias on negative binomial estimates, the study addresses estimation from datasets influenced by two types of event under-counting, and from disease transmission data subject to selection bias for successful outbreaks. CONCLUSIONS: Results show that maximum likelihood estimates of k can be biased upward by small sample size or under-reporting of zero-class events, but are not biased downward by any of the factors considered. Confidence intervals estimated from the asymptotic sampling variance tend to exhibit coverage below the nominal level, with overestimates of k comprising the great majority of coverage errors. Estimation from outbreak datasets does not increase the bias of k estimates, but can add significant upward bias to estimates of the mean. Because k varies inversely with the degree of overdispersion, these findings show that overestimation of the degree of overdispersion is very rare for these datasets.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

James O. Lloyd‐Smith (2007) studied this question.

synapsesocial.com/papers/6a6277bd5a5e300a328a0cb6https://doi.org/10.1371/journal.pone.0000180
Ask AI
Helpful
Bookmark
Share
View Full Paper