Using Web Search Query Data to Monitor Dengue Epidemics: A New Model for Neglected Tropical Disease Surveillance

被引:178
作者
Chan, Emily H. [1 ,2 ]
Sahai, Vikram [3 ]
Conrad, Corrie [3 ]
Brownstein, John S. [1 ,2 ,4 ]
机构
[1] Harvard Massachusetts Inst Technol, Div Hlth Sci & Technol, Childrens Hosp Informat Program, Boston, MA 02115 USA
[2] Childrens Hosp Boston, Div Emergency Med, Boston, MA USA
[3] Google Inc, Mountain View, CA USA
[4] Harvard Univ, Sch Med, Dept Pediat, Boston, MA 02115 USA
来源
PLOS NEGLECTED TROPICAL DISEASES | 2011年 / 5卷 / 05期
关键词
PREVENTION; VIRUS;
D O I
10.1371/journal.pntd.0001206
中图分类号
R51 [传染病];
学科分类号
100401 ;
摘要
Background: A variety of obstacles including bureaucracy and lack of resources have interfered with timely detection and reporting of dengue cases in many endemic countries. Surveillance efforts have turned to modern data sources, such as Internet search queries, which have been shown to be effective for monitoring influenza-like illnesses. However, few have evaluated the utility of web search query data for other diseases, especially those of high morbidity and mortality or where a vaccine may not exist. In this study, we aimed to assess whether web search queries are a viable data source for the early detection and monitoring of dengue epidemics. Methodology/Principal Findings: Bolivia, Brazil, India, Indonesia and Singapore were chosen for analysis based on available data and adequate search volume. For each country, a univariate linear model was then built by fitting a time series of the fraction of Google search query volume for specific dengue-related queries from that country against a time series of official dengue case counts for a time-frame within 2003-2010. The specific combination of queries used was chosen to maximize model fit. Spurious spikes in the data were also removed prior to model fitting. The final models, fit using a training subset of the data, were cross-validated against both the overall dataset and a holdout subset of the data. All models were found to fit the data quite well, with validation correlations ranging from 0.82 to 0.99. Conclusions/Significance: Web search query data were found to be capable of tracking dengue activity in Bolivia, Brazil, India, Indonesia and Singapore. Whereas traditional dengue data from official sources are often not available until after some substantial delay, web search query data are available in near real-time. These data represent valuable complement to assist with traditional dengue surveillance.
引用
收藏
页数:6
相关论文
共 24 条
  • [1] Best Practices in Dengue Surveillance: A Report from the Asia-Pacific and Americas Dengue Prevention Boards
    Beatty, Mark E.
    Stone, Amy
    Fitzsimons, David W.
    Hanna, Jeffrey N.
    Lam, Sai Kit
    Vong, Sirenda
    Guzman, Maria G.
    Mendez-Galvan, Jorge F.
    Halstead, Scott B.
    Letson, G. William
    Kuritsky, Joel
    Mahoney, Richard
    Margolis, Harold S.
    [J]. PLOS NEGLECTED TROPICAL DISEASES, 2010, 4 (11)
  • [2] Evaluation of school absenteeism data for early outbreak detection, New York City
    Besculides, M
    Heffernan, R
    Mostashari, F
    Weiss, D
    [J]. BMC PUBLIC HEALTH, 2005, 5 (1)
  • [3] Camacho Tania, 2004, Biomed., V24, P174
  • [4] Co-infections with Chikungunya Virus and Dengue Virus in Delhi, India
    Chahar, Harendra S.
    Bharaj, Preeti
    Dar, Lalit
    Guleria, Randeep
    Kabra, Sushill K.
    Broor, Shobha
    [J]. EMERGING INFECTIOUS DISEASES, 2009, 15 (07) : 1077 - 1080
  • [5] Hospital based clinical surveillance for dengue haemorrhagic fever in Bandung, Indonesia 1994-1995
    Chairulfatah, A
    Setiabudi, D
    Agoes, R
    van Sprundel, M
    Colebunders, R
    [J]. ACTA TROPICA, 2001, 80 (02) : 111 - 115
  • [6] Dengue and chikungunya infections in travelers
    Chen, Lin H.
    Wilson, Mary E.
    [J]. CURRENT OPINION IN INFECTIOUS DISEASES, 2010, 23 (05) : 438 - 444
  • [7] Choi H., 2009, Predicting the Present with Google Trends
  • [8] Timeliness of data sources used for influenza surveillance
    Dailey, Lynne
    Watkins, Rochelle E.
    Plant, Aileen J.
    [J]. JOURNAL OF THE AMERICAN MEDICAL INFORMATICS ASSOCIATION, 2007, 14 (05) : 626 - 631
  • [9] Das Debjani, 2005, MMWR Suppl, V54, P41
  • [10] Hospitalizations for suspected dengue in Puerto Rico, 1991-1995:: Estimation by capture-recapture methods
    Dechant, EJ
    Rigau-Pérez, JG
    [J]. AMERICAN JOURNAL OF TROPICAL MEDICINE AND HYGIENE, 1999, 61 (04) : 574 - 578