-
Deep learning models for predicting RNA degradation via dual crowdsourcing
Authors:
Hannah K. Wayment-Steele,
Wipapat Kladwang,
Andrew M. Watkins,
Do Soon Kim,
Bojan Tunguz,
Walter Reade,
Maggie Demkin,
Jonathan Romano,
Roger Wellington-Oguri,
John J. Nicol,
Jiayang Gao,
Kazuki Onodera,
Kazuki Fujikawa,
Hanfei Mao,
Gilles Vandewiele,
Michele Tinti,
Bram Steenwinckel,
Takuya Ito,
Taiga Noumi,
Shujun He,
Keiichiro Ishi,
Youhan Lee,
Fatih Öztürk,
Anthony Chiu,
Emin Öztürk
, et al. (4 additional authors not shown)
Abstract:
Messenger RNA-based medicines hold immense potential, as evidenced by their rapid deployment as COVID-19 vaccines. However, worldwide distribution of mRNA molecules has been limited by their thermostability, which is fundamentally limited by the intrinsic instability of RNA molecules to a chemical degradation reaction called in-line hydrolysis. Predicting the degradation of an RNA molecule is a ke…
▽ More
Messenger RNA-based medicines hold immense potential, as evidenced by their rapid deployment as COVID-19 vaccines. However, worldwide distribution of mRNA molecules has been limited by their thermostability, which is fundamentally limited by the intrinsic instability of RNA molecules to a chemical degradation reaction called in-line hydrolysis. Predicting the degradation of an RNA molecule is a key task in designing more stable RNA-based therapeutics. Here, we describe a crowdsourced machine learning competition ("Stanford OpenVaccine") on Kaggle, involving single-nucleotide resolution measurements on 6043 102-130-nucleotide diverse RNA constructs that were themselves solicited through crowdsourcing on the RNA design platform Eterna. The entire experiment was completed in less than 6 months, and 41% of nucleotide-level predictions from the winning model were within experimental error of the ground truth measurement. Furthermore, these models generalized to blindly predicting orthogonal degradation data on much longer mRNA molecules (504-1588 nucleotides) with improved accuracy compared to previously published models. Top teams integrated natural language processing architectures and data augmentation techniques with predictions from previous dynamic programming models for RNA secondary structure. These results indicate that such models are capable of representing in-line hydrolysis with excellent accuracy, supporting their use for designing stabilized messenger RNAs. The integration of two crowdsourcing platforms, one for data set creation and another for machine learning, may be fruitful for other urgent problems that demand scientific discovery on rapid timescales.
△ Less
Submitted 22 April, 2022; v1 submitted 14 October, 2021;
originally announced October 2021.
-
Religious Festivals and Influenza
Authors:
Alice P. Y. Chiu,
Qianying Lin,
Daihai He
Abstract:
Objectives Influenza outbreaks have been widely studied. However, the patterns between influenza and religious festivals remained unexplored. This study examined the patterns of influenza and Hanukkah in Israel, and that of influenza and Hajj in Bahrain, Egypt, Iraq, Jordan, Oman and Qatar. Method Influenza surveillance data of these seven countries from 2009 to 2017 were downloaded from the FluNe…
▽ More
Objectives Influenza outbreaks have been widely studied. However, the patterns between influenza and religious festivals remained unexplored. This study examined the patterns of influenza and Hanukkah in Israel, and that of influenza and Hajj in Bahrain, Egypt, Iraq, Jordan, Oman and Qatar. Method Influenza surveillance data of these seven countries from 2009 to 2017 were downloaded from the FluNet of the World Health Organization. Secondary data were collected for the countries' population, and the dates of Hajj and Hanukkah. We aggregated the weekly influenza A and B laboratory confirmations for each country over the study period. Weekly influenza A patterns and religious festival dates were further explored across the study period. Results We found that influenza A peaks closely followed Hanukkah in Israel in six out of seven years from 2010 to 2017. Aggregated influenza A peaks of the other six Middle East countries also occurred right after Hajj every year during the study period. Conclusions We predict that unless there is an emergence of new influenza strain, such influenza patterns are likely to persist in future years. Our results suggested that the optimal timing of mass influenza vaccination should take into considerations of the dates of these religious festivals.
△ Less
Submitted 24 October, 2017;
originally announced October 2017.
-
Patterns of Influenza Vaccination Coverage in the United States from 2009 to 2015
Authors:
Alice P. Y. Chiu,
Duo Yu,
Jonathan Dushoff,
Daihai He
Abstract:
Background: Globally, influenza is a major cause of morbidity, hospitalization and mortality. Influenza vaccination has shown substantial protective effectiveness in the United States. We investigated state-level patterns of coverage rates of seasonal and pandemic influenza vaccination, among the overall population in the U.S. and specifically among children and the elderly, from 2009/10 to 2014/1…
▽ More
Background: Globally, influenza is a major cause of morbidity, hospitalization and mortality. Influenza vaccination has shown substantial protective effectiveness in the United States. We investigated state-level patterns of coverage rates of seasonal and pandemic influenza vaccination, among the overall population in the U.S. and specifically among children and the elderly, from 2009/10 to 2014/15, and associations with ecological factors.
Methods and Findings: We obtained state-level influenza vaccination coverage rates from national surveys, and state-level socio-demographic and health data from a variety of sources. We employed a retrospective ecological study design, and used mixed-model regression to determine the levels of ecological association of the state-level vaccinations rates with these factors, both with and without region as a factor for the three populations. We found that health-care access is positively and significantly associated with mean influenza vaccination coverage rates across all populations and models. We also found that prevalence of asthma in adults are negatively and significantly associated with mean influenza vaccination coverage rates in the elderly populations.
Conclusions: Health-care access has a robust, positive association with state-level vaccination rates across different populations. This highlights a potential population-level advantage of expanding health-care access.
△ Less
Submitted 13 March, 2017;
originally announced March 2017.
-
Increasing Trends of Guillain-Barré Syndrome (GBS) and Dengue in Hong Kong
Authors:
Xiujuan Tang,
Shi Zhao,
Alice P. Y. Chiu,
Xin Wang,
Lin Yang,
Daihai He
Abstract:
Background: Guillain-Barré Syndrome (GBS) is a common type of severe acute paralytic neuropathy and associated with other virus infections such as dengue fever and Zika. This study investigate the relationship between GBS, dengue, local meteorological factors in Hong Kong and global climatic factors from January 2000 to June 2016.
Methods: The correlations between GBS, dengue, Multivariate El Ni…
▽ More
Background: Guillain-Barré Syndrome (GBS) is a common type of severe acute paralytic neuropathy and associated with other virus infections such as dengue fever and Zika. This study investigate the relationship between GBS, dengue, local meteorological factors in Hong Kong and global climatic factors from January 2000 to June 2016.
Methods: The correlations between GBS, dengue, Multivariate El Nino Southern Oscillation Index (MEI) and local meteorological data were explored by the Spearman Rank correlations and cross-correlations between these time series. Poisson regression models were fitted to identify nonlinear associations between MEI and dengue. Cross wavelet analysis was applied to infer potential non-stationary oscillating associations among MEI, dengue and GBS.
Findings : An increasing trend was found for both GBS cases and imported dengue cases in Hong Kong. We found a weak but statistically significant negative correlation between GBS and local meteorological factors. MEI explained over 12\% of dengue's variations from Poisson regression models. Wavelet analyses showed that there is possible non-stationary oscillating association between dengue and GBS from 2005 to 2015 in Hong Kong. Our study has led to an improved understanding of the timing and relationship between GBS, dengue and MEI.
△ Less
Submitted 13 March, 2017;
originally announced March 2017.
-
Effects of Reactive Social Distancing on the 1918 Influenza Pandemic
Authors:
Duo Yu,
Qianying Lin,
Alice PY Chiu,
Daihai He
Abstract:
The 1918 influenza pandemic was characterized by multiple epidemic waves. We investigated into reactive social distancing, a form of behavioral responses, and its effect on the multiple influenza waves in the United Kingdom. Two forms of reactive social distancing have been used in previous studies: Power function, which is a function of the proportion of recent influenza mortality in a population…
▽ More
The 1918 influenza pandemic was characterized by multiple epidemic waves. We investigated into reactive social distancing, a form of behavioral responses, and its effect on the multiple influenza waves in the United Kingdom. Two forms of reactive social distancing have been used in previous studies: Power function, which is a function of the proportion of recent influenza mortality in a population, and Hill function, which is a function of the actual number of recent influenza mortality. Using a simple epidemic model with a Power function and one common set of parameters, we provided a good model fit for the observed multiple epidemic waves in London boroughs, Birmingham and Liverpool. Our approach is different from previous studies where separate models are fitted to each city. We then applied these model parameters obtained from fitting three cities to all 334 administrative units in England and Wales and including the population sizes of individual administrative units. We computed the Pearson's correlation between the observed and simulated data for each administrative unit. We achieved a median correlation of 0.636, indicating our model predictions perform reasonably well. Our modelling approach which requires reduced number of parameters resulted in computational efficiency gain without over-fitting the model. Our works have both scientific and public health significance.
△ Less
Submitted 12 March, 2017;
originally announced March 2017.
-
Spatio-temporal patterns of influenza B proportions
Authors:
Daihai He,
Alice PY Chiu,
Qianying Lin,
Duo Yu
Abstract:
We study the spatio-temporal patterns of the proportion of influenza B out of laboratory confirmations of both influenza A and B, with data from 139 countries and regions downloaded from the FluNet compiled by the World Health Organization, from January 2006 to October 2015, excluding 2009. We restricted our analysis to 34 countries that reported more than 2000 confirmations for each of types A an…
▽ More
We study the spatio-temporal patterns of the proportion of influenza B out of laboratory confirmations of both influenza A and B, with data from 139 countries and regions downloaded from the FluNet compiled by the World Health Organization, from January 2006 to October 2015, excluding 2009. We restricted our analysis to 34 countries that reported more than 2000 confirmations for each of types A and B over the study period. We find that Pearson's correlation is 0.669 between effective distance from Mexico and influenza B proportion among the countries from January 2006 to October 2015. In the United States, influenza B proportion in the pre-pandemic period (2003-2008) negatively correlated with that in the post-pandemic era (2010-2015) at the regional level. Our study limitations are the country-level variations in both surveillance methods and testing policies. Influenza B proportion displayed wide variations over the study period. Our findings suggest that even after excluding 2009's data, the influenza pandemic still has an evident impact on the relative burden of the two influenza types. Future studies could examine whether there are other additional factors. This study has potential implications in prioritizing public health control measures.
△ Less
Submitted 26 January, 2016;
originally announced January 2016.