Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Aishwarya Sarkar

Enhanced Soups for Graph Neural Networks

Mar 14, 2025

Joseph Zuber, Aishwarya Sarkar, Joseph Jennings, Ali Jannesari

Abstract:Graph Neural Networks (GNN) have demonstrated state-of-the-art performance in numerous scientific and high-performance computing (HPC) applications. Recent work suggests that "souping" (combining) individually trained GNNs into a single model can improve performance without increasing compute and memory costs during inference. However, existing souping algorithms are often slow and memory-intensive, which limits their scalability. We introduce Learned Souping for GNNs, a gradient-descent-based souping strategy that substantially reduces time and memory overhead compared to existing methods. Our approach is evaluated across multiple Open Graph Benchmark (OGB) datasets and GNN architectures, achieving up to 1.2% accuracy improvement and 2.1X speedup. Additionally, we propose Partition Learned Souping, a novel partition-based variant of learned souping that significantly reduces memory usage. On the ogbn-products dataset with GraphSAGE, partition learned souping achieves a 24.5X speedup and a 76% memory reduction without compromising accuracy.

* 10 pages, 4 figures, 3 tables, accepted to GrAPL 2025 (colocated with IPDPS 2025)

Via

Access Paper or Ask Questions

MassiveGNN: Efficient Training via Prefetching for Massively Connected Distributed Graphs

Oct 30, 2024

Aishwarya Sarkar, Sayan Ghosh, Nathan R. Tallent, Ali Jannesari

Abstract:Graph Neural Networks (GNN) are indispensable in learning from graph-structured data, yet their rising computational costs, especially on massively connected graphs, pose significant challenges in terms of execution performance. To tackle this, distributed-memory solutions such as partitioning the graph to concurrently train multiple replicas of GNNs are in practice. However, approaches requiring a partitioned graph usually suffer from communication overhead and load imbalance, even under optimal partitioning and communication strategies due to irregularities in the neighborhood minibatch sampling. This paper proposes practical trade-offs for improving the sampling and communication overheads for representation learning on distributed graphs (using popular GraphSAGE architecture) by developing a parameterized continuous prefetch and eviction scheme on top of the state-of-the-art Amazon DistDGL distributed GNN framework, demonstrating about 15-40% improvement in end-to-end training performance on the National Energy Research Scientific Computing Center's (NERSC) Perlmutter supercomputer for various OGB datasets.

* In Proc. of the IEEE International Conference on Cluster Computing (CLUSTER), 2024

Via

Access Paper or Ask Questions

The Landscape and Challenges of HPC Research and LLMs

Feb 07, 2024

Le Chen, Nesreen K. Ahmed, Akash Dutta, Arijit Bhattacharjee, Sixing Yu, Quazi Ishtiaque Mahmud, Waqwoya Abebe, Hung Phan, Aishwarya Sarkar, Branden Butler(+7 more)

Figure 1 for The Landscape and Challenges of HPC Research and LLMs

Figure 2 for The Landscape and Challenges of HPC Research and LLMs

Abstract:Recently, language models (LMs), especially large language models (LLMs), have revolutionized the field of deep learning. Both encoder-decoder models and prompt-based techniques have shown immense potential for natural language processing and code-based tasks. Over the past several years, many research labs and institutions have invested heavily in high-performance computing, approaching or breaching exascale performance levels. In this paper, we posit that adapting and utilizing such language model-based techniques for tasks in high-performance computing (HPC) would be very beneficial. This study presents our reasoning behind the aforementioned position and highlights how existing ideas can be improved and adapted for HPC tasks.

Via

Access Paper or Ask Questions

Accelerating Domain-aware Deep Learning Models with Distributed Training

Jan 25, 2023

Aishwarya Sarkar, Chaoqun Lu, Ali Jannesari

Abstract:Recent advances in data-generating techniques led to an explosive growth of geo-spatiotemporal data. In domains such as hydrology, ecology, and transportation, interpreting the complex underlying patterns of spatiotemporal interactions with the help of deep learning techniques hence becomes the need of the hour. However, applying deep learning techniques without domain-specific knowledge tends to provide sub-optimal prediction performance. Secondly, training such models on large-scale data requires extensive computational resources. To eliminate these challenges, we present a novel distributed domain-aware spatiotemporal network that utilizes domain-specific knowledge with improved model performance. Our network consists of a pixel-contribution block, a distributed multiheaded multichannel convolutional (CNN) spatial block, and a recurrent temporal block. We choose flood prediction in hydrology as a use case to test our proposed method. From our analysis, the network effectively predicts high peaks in discharge measurements at watershed outlets with up to 4.1x speedup and increased prediction performance of up to 93\%. Our approach achieved a 12.6x overall speedup and increased the mean prediction performance by 16\%. We perform extensive experiments on a dataset of 23 watersheds in a northern state of the U.S. and present our findings.

* Accepted for Workshop on Multi-scale, Multi-physic and Coupled Problems on Highly Parallel Systems, HPC Asia 2023, 27 February - 2 March 2023, Singapore

Via

Access Paper or Ask Questions

Transfer Learning Approaches for Knowledge Discovery in Grid-based Geo-Spatiotemporal Data

Oct 02, 2021

Aishwarya Sarkar, Jien Zhang, Chaoqun Lu, Ali Jannesari

Figure 1 for Transfer Learning Approaches for Knowledge Discovery in Grid-based Geo-Spatiotemporal Data

Figure 2 for Transfer Learning Approaches for Knowledge Discovery in Grid-based Geo-Spatiotemporal Data

Abstract:Extracting and meticulously analyzing geo-spatiotemporal features is crucial to recognize intricate underlying causes of natural events, such as floods. Limited evidence about hidden factors leading to climate change makes it challenging to predict regional water discharge accurately. In addition, the explosive growth in complex geo-spatiotemporal environment data that requires repeated learning by the state-of-the-art neural networks for every new region emphasizes the need for new computationally efficient methods, advanced computational resources, and extensive training on a massive amount of available monitored data. We, therefore, propose HydroDeep, an effectively reusable pretrained model to address this problem of transferring knowledge from one region to another by effectively capturing their intrinsic geo-spatiotemporal variance. Further, we present four transfer learning approaches on HydroDeep for spatiotemporal interpretability that improve Nash-Sutcliffe efficiency by 9% to 108% in new regions with a 95% reduction in time.

Via

Access Paper or Ask Questions

HydroDeep -- A Knowledge Guided Deep Neural Network for Geo-Spatiotemporal Data Analysis

Oct 09, 2020

Aishwarya Sarkar, Jien Zhang, Chaoqun Lu, Ali Jannesari

Figure 1 for HydroDeep -- A Knowledge Guided Deep Neural Network for Geo-Spatiotemporal Data Analysis

Figure 2 for HydroDeep -- A Knowledge Guided Deep Neural Network for Geo-Spatiotemporal Data Analysis

Figure 3 for HydroDeep -- A Knowledge Guided Deep Neural Network for Geo-Spatiotemporal Data Analysis

Figure 4 for HydroDeep -- A Knowledge Guided Deep Neural Network for Geo-Spatiotemporal Data Analysis

Abstract:Floods are one of the major climate-related disasters, leading to substantial economic loss and social safety issue. However, the confidence in predicting changes in fluvial floods remains low due to limited evidence and complex causes of regional climate change. The recent development in machine learning techniques has the potential to improve traditional hydrological models by using monitoring data. Although Recurrent Neural Networks (RNN) perform remarkably with multivariate time series data, these models are blinded to the underlying mechanisms represented in a process-based model for flood prediction. While both process-based models and deep learning networks have their strength, understanding the fundamental mechanisms intrinsic to geo-spatiotemporal information is crucial to improve the prediction accuracy of flood occurrence. This paper demonstrates a neural network architecture (HydroDeep) that couples a process-based hydro-ecological model with a combination of Deep Convolutional Neural Network (CNN) and Long Short-Term Memory (LSTM) Network to build a hybrid baseline model. HydroDeep outperforms the performance of both the independent networks by 4.8% and 31.8% respectively in Nash-Sutcliffe efficiency. A trained HydroDeep can transfer its knowledge and can learn the Geo-spatiotemporal features of any new region in minimal training iterations.

Via

Access Paper or Ask Questions