IDEAS home Printed from https://ideas.repec.org/a/plo/pone00/0261629.html
   My bibliography  Save this article

Analysis and prediction of hand, foot and mouth disease incidence in China using Random Forest and XGBoost

Author

Listed:
  • Delin Meng
  • Jun Xu
  • Jijun Zhao

Abstract

Hand, foot and mouth disease (HFMD) is an increasingly serious public health problem, and it has caused an outbreak in China every year since 2008. Predicting the incidence of HFMD and analyzing its influential factors are of great significance to its prevention. Now, machine learning has shown advantages in infectious disease models, but there are few studies on HFMD incidence based on machine learning that cover all the provinces in mainland China. In this study, we proposed two different machine learning algorithms, Random Forest and eXtreme Gradient Boosting (XGBoost), to perform our analysis and prediction. We first used Random Forest to examine the association between HFMD incidence and potential influential factors for 31 provinces in mainland China. Next, we established Random Forest and XGBoost prediction models using meteorological and social factors as the predictors. Finally, we applied our prediction models in four different regions of mainland China and evaluated the performance of them. Our results show that: 1) Meteorological factors and social factors jointly affect the incidence of HFMD in mainland China. Average temperature and population density are the two most significant influential factors; 2) Population flux has different delayed effect in affecting HFMD incidence in different regions. From a national perspective, the model using population flux data delayed for one month has better prediction performance; 3) The prediction capability of XGBoost model was better than that of Random Forest model from the overall perspective. XGBoost model is more suitable for predicting the incidence of HFMD in mainland China.

Suggested Citation

  • Delin Meng & Jun Xu & Jijun Zhao, 2021. "Analysis and prediction of hand, foot and mouth disease incidence in China using Random Forest and XGBoost," PLOS ONE, Public Library of Science, vol. 16(12), pages 1-16, December.
  • Handle: RePEc:plo:pone00:0261629
    DOI: 10.1371/journal.pone.0261629
    as

    Download full text from publisher

    File URL: https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0261629
    Download Restriction: no

    File URL: https://journals.plos.org/plosone/article/file?id=10.1371/journal.pone.0261629&type=printable
    Download Restriction: no

    File URL: https://libkey.io/10.1371/journal.pone.0261629?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Matthew J. Ferrari & Rebecca F. Grais & Nita Bharti & Andrew J. K. Conlan & Ottar N. Bjørnstad & Lara J. Wolfson & Philippe J. Guerin & Ali Djibo & Bryan T. Grenfell, 2008. "The dynamics of measles in sub-Saharan Africa," Nature, Nature, vol. 451(7179), pages 679-684, February.
    2. Malki, Zohair & Atlam, El-Sayed & Hassanien, Aboul Ella & Dagnew, Guesh & Elhosseini, Mostafa A. & Gad, Ibrahim, 2020. "Association between weather data and COVID-19 pandemic predicting mortality rate: Machine learning approaches," Chaos, Solitons & Fractals, Elsevier, vol. 138(C).
    3. Huifen Feng & Guangcai Duan & Rongguang Zhang & Weidong Zhang, 2014. "Time Series Analysis of Hand-Foot-Mouth Disease Hospitalization in Zhengzhou: Establishment of Forecasting Models Using Climate Variables as Predictors," PLOS ONE, Public Library of Science, vol. 9(1), pages 1-10, January.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Wayne M. Getz & Jean-Paul Gonzalez & Richard Salter & James Bangura & Colin Carlson & Moinya Coomber & Eric Dougherty & David Kargbo & Nathan D. Wolfe & Nadia Wauquier, 2015. "Tactics and Strategies for Managing Ebola Outbreaks and the Salience of Immunization," Post-Print hal-01214432, HAL.
    2. Das, Ayan Kumar & Kalam, Sidra & Kumar, Chiranjeev & Sinha, Ditipriya, 2021. "TLCoV- An automated Covid-19 screening model using Transfer Learning from chest X-ray images," Chaos, Solitons & Fractals, Elsevier, vol. 144(C).
    3. Vivek Jason Jayaraj & Victor Chee Wai Hoe, 2022. "Forecasting HFMD Cases Using Weather Variables and Google Search Queries in Sabah, Malaysia," IJERPH, MDPI, vol. 19(24), pages 1-9, December.
    4. Jijun Zhao & Xinmin Li, 2016. "Determinants of the Transmission Variation of Hand, Foot and Mouth Disease in China," PLOS ONE, Public Library of Science, vol. 11(10), pages 1-14, October.
    5. Wang, Mingzhao & Fu, Zuntao, 2022. "A new method of nonlinear causality detection: Reservoir computing Granger causality," Chaos, Solitons & Fractals, Elsevier, vol. 154(C).
    6. Yu-Tse Tsan & Endah Kristiani & Po-Yu Liu & Wei-Min Chu & Chao-Tung Yang, 2022. "In the Seeking of Association between Air Pollutant and COVID-19 Confirmed Cases Using Deep Learning," IJERPH, MDPI, vol. 19(11), pages 1-19, May.
    7. Pedro Henrique Melo Albuquerque & Yaohao Peng & João Pedro Fontoura da Silva, 2022. "Making the whole greater than the sum of its parts: A literature review of ensemble methods for financial time series forecasting," Journal of Forecasting, John Wiley & Sons, Ltd., vol. 41(8), pages 1701-1724, December.
    8. Rasheed, Jawad & Jamil, Akhtar & Hameed, Alaa Ali & Aftab, Usman & Aftab, Javaria & Shah, Syed Attique & Draheim, Dirk, 2020. "A survey on artificial intelligence approaches in supporting frontline workers and decision makers for the COVID-19 pandemic," Chaos, Solitons & Fractals, Elsevier, vol. 141(C).
    9. Alexander D Becker & Bryan T Grenfell, 2017. "tsiR: An R package for time-series Susceptible-Infected-Recovered models of epidemics," PLOS ONE, Public Library of Science, vol. 12(9), pages 1-10, September.
    10. Wan Yang & Liang Wen & Shen-Long Li & Kai Chen & Wen-Yi Zhang & Jeffrey Shaman, 2017. "Geospatial characteristics of measles transmission in China during 2005−2014," PLOS Computational Biology, Public Library of Science, vol. 13(4), pages 1-21, April.
    11. van de Water, Antoinette & Henley, Michelle & Bates, Lucy & Slotow, Rob, 2022. "The value of elephants: A pluralist approach," Ecosystem Services, Elsevier, vol. 58(C).
    12. Daniel M. Parker & James W. Wood & Shinsuke Tomita & Sharon DeWitte & Julia Jennings & Liwang Cui, 2014. "Household ecology and out-migration among ethnic Karen along the Thai-Myanmar border," Demographic Research, Max Planck Institute for Demographic Research, Rostock, Germany, vol. 30(39), pages 1129-1156.
    13. Frederik Seeup Hass & Jamal Jokar Arsanjani, 2021. "The Geography of the Covid-19 Pandemic: A Data-Driven Approach to Exploring Geographical Driving Forces," IJERPH, MDPI, vol. 18(6), pages 1-19, March.
    14. Francis Tuluri & Reddy Remata & Wilbur L. Walters & Paul. B. Tchounwou, 2022. "Application of Machine Learning to Study the Association between Environmental Factors and COVID-19 Cases in Mississippi, USA," Mathematics, MDPI, vol. 10(6), pages 1-9, March.
    15. Pang, Liuyong & Ruan, Shigui & Liu, Sanhong & Zhao, Zhong & Zhang, Xinan, 2015. "Transmission dynamics and optimal control of measles epidemics," Applied Mathematics and Computation, Elsevier, vol. 256(C), pages 131-147.
    16. Rizvi, Syeda Amna & Umair, Muhammad & Cheema, Muhammad Aamir, 2021. "Clustering of countries for COVID-19 cases based on disease prevalence, health systems and environmental indicators," Chaos, Solitons & Fractals, Elsevier, vol. 151(C).
    17. Tayarani N., Mohammad-H., 2021. "Applications of artificial intelligence in battling against covid-19: A literature review," Chaos, Solitons & Fractals, Elsevier, vol. 142(C).

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:plo:pone00:0261629. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: plosone (email available below). General contact details of provider: https://journals.plos.org/plosone/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.