Respiratory syncytial virus tracking using internet search engine data
AffiliationUniv Arizona, Coll Publ Hlth, Div Epidemiol & Biostat, Tucson, AZ 85721 USA
MetadataShow full item record
PublisherBIOMED CENTRAL LTD
CitationOren et al. BMC Public Health (2018) 18:445 https://doi.org/10.1186/s12889-018-5367-z
JournalBMC PUBLIC HEALTH
Rights© The Author(s). 2018 Open Access This article is distributed under the terms of the Creative Commons Attribution 4.0 International License.
Collection InformationThis item from the UA Faculty Publications collection is made available by the University of Arizona with support from the University of Arizona Libraries. If you have questions, please contact us at firstname.lastname@example.org.
AbstractBackground: Respiratory Syncytial Virus (RSV) is the leading cause of hospitalization in children less than 1 year of age in the United States. Internet search engine queries may provide high resolution temporal and spatial data to estimate and predict disease activity. Methods: After filtering an initial list of 613 symptoms using high-resolution Bing search logs, we used Google Trends data between 2004 and 2016 for a smaller list of 50 terms to build predictive models of RSV incidence for five states where long-term surveillance data was available. We then used domain adaptation to model RSV incidence for the 45 remaining US states. Results: Surveillance data sources (hospitalization and laboratory reports) were highly correlated, as were laboratory reports with search engine data. The four terms which were most often statistically significantly correlated as time series with the surveillance data in the five state models were RSV, flu, pneumonia, and bronchiolitis. Using our models, we tracked the spread of RSV by observing the time of peak use of the search term in different states. In general, the RSV peak moved from south-east (Florida) to the north-west US. Conclusions: Our study represents the first time that RSV has been tracked using Internet data results and highlights successful use of search filters and domain adaptation techniques, using data at multiple resolutions. Our approach may assist in identifying spread of both local and more widespread RSV transmission and may be applicable to other seasonal conditions where comprehensive epidemiological data is difficult to collect or obtain.
NoteOpen access journal.
VersionFinal published version
- Correlation between respiratory syncytial virus (RSV) test data and hospitalization of children for RSV lower respiratory tract illness in Florida.
- Authors: Light M, Bauman J, Mavunda K, Malinoski F, Eggleston M
- Issue date: 2008 Jun
- Association between respiratory syncytial virus activity and pneumococcal disease in infants: a time series analysis of US hospitalization data.
- Authors: Weinberger DM, Klugman KP, Steiner CA, Simonsen L, Viboud C
- Issue date: 2015 Jan
- Respiratory syncytial virus activity-- United States, July 2007-December 2008.
- Authors: Centers for Disease Control and Prevention (CDC).
- Issue date: 2008 Dec 19
- Incidence and clinical features of respiratory syncytial virus infections in a population-based surveillance site in the Nile Delta Region.
- Authors: Rowlinson E, Dueger E, Taylor T, Mansour A, Van Beneden C, Abukela M, Zhang X, Refaey S, Bastawy H, Kandeel A
- Issue date: 2013 Dec 15
- Substantial variability in community respiratory syncytial virus season timing.
- Authors: Mullins JA, Lamonte AC, Bresee JS, Anderson LJ
- Issue date: 2003 Oct