Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Laura Natali

Neural Network Training with Highly Incomplete Datasets

Jul 01, 2021

Yu-Wei Chang, Laura Natali, Oveis Jamialahmadi, Stefano Romeo, Joana B. Pereira, Giovanni Volpe

Figure 1 for Neural Network Training with Highly Incomplete Datasets

Figure 2 for Neural Network Training with Highly Incomplete Datasets

Figure 3 for Neural Network Training with Highly Incomplete Datasets

Figure 4 for Neural Network Training with Highly Incomplete Datasets

Abstract:Neural network training and validation rely on the availability of large high-quality datasets. However, in many cases only incomplete datasets are available, particularly in health care applications, where each patient typically undergoes different clinical procedures or can drop out of a study. Since the data to train the neural networks need to be complete, most studies discard the incomplete datapoints, which reduces the size of the training data, or impute the missing features, which can lead to artefacts. Alas, both approaches are inadequate when a large portion of the data is missing. Here, we introduce GapNet, an alternative deep-learning training approach that can use highly incomplete datasets. First, the dataset is split into subsets of samples containing all values for a certain cluster of features. Then, these subsets are used to train individual neural networks. Finally, this ensemble of neural networks is combined into a single neural network whose training is fine-tuned using all complete datapoints. Using two highly incomplete real-world medical datasets, we show that GapNet improves the identification of patients with underlying Alzheimer's disease pathology and of patients at risk of hospitalization due to Covid-19. By distilling the information available in incomplete datasets without having to reduce their size or to impute missing values, GapNet will permit to extract valuable information from a wide range of datasets, benefiting diverse fields from medicine to engineering.

* 11 pages, 3 figures, 1 table

Via

Access Paper or Ask Questions

Improving epidemic testing and containment strategies using machine learning

Nov 23, 2020

Laura Natali, Saga Helgadottir, Onofrio M. Marago, Giovanni Volpe

Figure 1 for Improving epidemic testing and containment strategies using machine learning

Figure 2 for Improving epidemic testing and containment strategies using machine learning

Figure 3 for Improving epidemic testing and containment strategies using machine learning

Figure 4 for Improving epidemic testing and containment strategies using machine learning

Abstract:Containment of epidemic outbreaks entails great societal and economic costs. Cost-effective containment strategies rely on efficiently identifying infected individuals, making the best possible use of the available testing resources. Therefore, quickly identifying the optimal testing strategy is of critical importance. Here, we demonstrate that machine learning can be used to identify which individuals are most beneficial to test, automatically and dynamically adapting the testing strategy to the characteristics of the disease outbreak. Specifically, we simulate an outbreak using the archetypal susceptible-infectious-recovered (SIR) model and we use data about the first confirmed cases to train a neural network that learns to make predictions about the rest of the population. Using these prediction, we manage to contain the outbreak more effectively and more quickly than with standard approaches. Furthermore, we demonstrate how this method can be used also when there is a possibility of reinfection (SIRS model) to efficiently eradicate an endemic disease.

* 11 pages, 4 figures

Via

Access Paper or Ask Questions