Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Andreas Hellander

Exploiting the Asymmetric Uncertainty Structure of Pre-trained VLMs on the Unit Hypersphere

May 16, 2025

Li Ju, Max Andersson, Stina Fredriksson, Edward Glöckner, Andreas Hellander, Ekta Vats, Prashant Singh

Abstract:Vision-language models (VLMs) as foundation models have significantly enhanced performance across a wide range of visual and textual tasks, without requiring large-scale training from scratch for downstream tasks. However, these deterministic VLMs fail to capture the inherent ambiguity and uncertainty in natural language and visual data. Recent probabilistic post-hoc adaptation methods address this by mapping deterministic embeddings onto probability distributions; however, existing approaches do not account for the asymmetric uncertainty structure of the modalities, and the constraint that meaningful deterministic embeddings reside on a unit hypersphere, potentially leading to suboptimal performance. In this paper, we address the asymmetric uncertainty structure inherent in textual and visual data, and propose AsymVLM to build probabilistic embeddings from pre-trained VLMs on the unit hypersphere, enabling uncertainty quantification. We validate the effectiveness of the probabilistic embeddings on established benchmarks, and present comprehensive ablation studies demonstrating the inherent nature of asymmetry in the uncertainty structure of textual and visual data.

Via

Access Paper or Ask Questions

ConDiSim: Conditional Diffusion Models for Simulation Based Inference

May 13, 2025

Mayank Nautiyal, Andreas Hellander, Prashant Singh

Abstract:We present a conditional diffusion model - ConDiSim, for simulation-based inference of complex systems with intractable likelihoods. ConDiSim leverages denoising diffusion probabilistic models to approximate posterior distributions, consisting of a forward process that adds Gaussian noise to parameters, and a reverse process learning to denoise, conditioned on observed data. This approach effectively captures complex dependencies and multi-modalities within posteriors. ConDiSim is evaluated across ten benchmark problems and two real-world test problems, where it demonstrates effective posterior approximation accuracy while maintaining computational efficiency and stability in model training. ConDiSim offers a robust and extensible framework for simulation-based inference, particularly suitable for parameter inference workflows requiring fast inference methods.

Via

Access Paper or Ask Questions

Variational Autoencoders for Efficient Simulation-Based Inference

Nov 21, 2024

Mayank Nautiyal, Andrey Shternshis, Andreas Hellander, Prashant Singh

Figure 1 for Variational Autoencoders for Efficient Simulation-Based Inference

Figure 2 for Variational Autoencoders for Efficient Simulation-Based Inference

Figure 3 for Variational Autoencoders for Efficient Simulation-Based Inference

Figure 4 for Variational Autoencoders for Efficient Simulation-Based Inference

Abstract:We present a generative modeling approach based on the variational inference framework for likelihood-free simulation-based inference. The method leverages latent variables within variational autoencoders to efficiently estimate complex posterior distributions arising from stochastic simulations. We explore two variations of this approach distinguished by their treatment of the prior distribution. The first model adapts the prior based on observed data using a multivariate prior network, enhancing generalization across various posterior queries. In contrast, the second model utilizes a standard Gaussian prior, offering simplicity while still effectively capturing complex posterior distributions. We demonstrate the efficacy of these models on well-established benchmark problems, achieving results comparable to flow-based approaches while maintaining computational efficiency and scalability.

Via

Access Paper or Ask Questions

Toward efficient resource utilization at edge nodes in federated learning

Sep 19, 2023

Sadi Alawadi, Addi Ait-Mlouk, Salman Toor, Andreas Hellander

Abstract:Federated learning (FL) enables edge nodes to collaboratively contribute to constructing a global model without sharing their data. This is accomplished by devices computing local, private model updates that are then aggregated by a server. However, computational resource constraints and network communication can become a severe bottleneck for larger model sizes typical for deep learning applications. Edge nodes tend to have limited hardware resources (RAM, CPU), and the network bandwidth and reliability at the edge is a concern for scaling federated fleet applications. In this paper, we propose and evaluate a FL strategy inspired by transfer learning in order to reduce resource utilization on devices, as well as the load on the server and network in each global training round. For each local model update, we randomly select layers to train, freezing the remaining part of the model. In doing so, we can reduce both server load and communication costs per round by excluding all untrained layer weights from being transferred to the server. The goal of this study is to empirically explore the potential trade-off between resource utilization on devices and global model convergence under the proposed strategy. We implement the approach using the federated learning framework FEDn. A number of experiments were carried out over different datasets (CIFAR-10, CASA, and IMDB), performing different tasks using different deep-learning model architectures. Our results show that training the model partially can accelerate the training process, efficiently utilizes resources on-device, and reduce the data transmission by around 75% and 53% when we train 25%, and 50% of the model layers, respectively, without harming the resulting global model accuracy.

* 16 pages, 5 tables, 8 figures

Via

Access Paper or Ask Questions

FedBot: Enhancing Privacy in Chatbots with Federated Learning

Apr 04, 2023

Addi Ait-Mlouk, Sadi Alawadi, Salman Toor, Andreas Hellander

Abstract:Chatbots are mainly data-driven and usually based on utterances that might be sensitive. However, training deep learning models on shared data can violate user privacy. Such issues have commonly existed in chatbots since their inception. In the literature, there have been many approaches to deal with privacy, such as differential privacy and secure multi-party computation, but most of them need to have access to users' data. In this context, Federated Learning (FL) aims to protect data privacy through distributed learning methods that keep the data in its location. This paper presents Fedbot, a proof-of-concept (POC) privacy-preserving chatbot that leverages large-scale customer support data. The POC combines Deep Bidirectional Transformer models and federated learning algorithms to protect customer data privacy during collaborative model training. The results of the proof-of-concept showcase the potential for privacy-preserving chatbots to transform the customer support industry by delivering personalized and efficient customer service that meets data privacy regulations and legal requirements. Furthermore, the system is specifically designed to improve its performance and accuracy over time by leveraging its ability to learn from previous interactions.

Via

Access Paper or Ask Questions

Accelerating Fair Federated Learning: Adaptive Federated Adam

Jan 23, 2023

Li Ju, Tianru Zhang, Salman Toor, Andreas Hellander

Figure 1 for Accelerating Fair Federated Learning: Adaptive Federated Adam

Figure 2 for Accelerating Fair Federated Learning: Adaptive Federated Adam

Figure 3 for Accelerating Fair Federated Learning: Adaptive Federated Adam

Figure 4 for Accelerating Fair Federated Learning: Adaptive Federated Adam

Abstract:Federated learning is a distributed and privacy-preserving approach to train a statistical model collaboratively from decentralized data of different parties. However, when datasets of participants are not independent and identically distributed (non-IID), models trained by naive federated algorithms may be biased towards certain participants, and model performance across participants is non-uniform. This is known as the fairness problem in federated learning. In this paper, we formulate fairness-controlled federated learning as a dynamical multi-objective optimization problem to ensure fair performance across all participants. To solve the problem efficiently, we study the convergence and bias of Adam as the server optimizer in federated learning, and propose Adaptive Federated Adam (AdaFedAdam) to accelerate fair federated learning with alleviated bias. We validated the effectiveness, Pareto optimality and robustness of AdaFedAdam in numerical experiments and show that AdaFedAdam outperforms existing algorithms, providing better convergence and fairness properties of the federated scheme.

Via

Access Paper or Ask Questions

FedQAS: Privacy-aware machine reading comprehension with federated learning

Feb 09, 2022

Addi Ait-Mlouk, Sadi Alawadi, Salman Toor, Andreas Hellander

Figure 1 for FedQAS: Privacy-aware machine reading comprehension with federated learning

Figure 2 for FedQAS: Privacy-aware machine reading comprehension with federated learning

Figure 3 for FedQAS: Privacy-aware machine reading comprehension with federated learning

Figure 4 for FedQAS: Privacy-aware machine reading comprehension with federated learning

Abstract:Machine reading comprehension (MRC) of text data is one important task in Natural Language Understanding. It is a complex NLP problem with a lot of ongoing research fueled by the release of the Stanford Question Answering Dataset (SQuAD) and Conversational Question Answering (CoQA). It is considered to be an effort to teach computers how to "understand" a text, and then to be able to answer questions about it using deep learning. However, until now large-scale training on private text data and knowledge sharing has been missing for this NLP task. Hence, we present FedQAS, a privacy-preserving machine reading system capable of leveraging large-scale private data without the need to pool those datasets in a central location. The proposed approach combines transformer models and federated learning technologies. The system is developed using the FEDn framework and deployed as a proof-of-concept alliance initiative. FedQAS is flexible, language-agnostic, and allows intuitive participation and execution of local model training. In addition, we present the architecture and implementation of the system, as well as provide a reference evaluation based on the SQUAD dataset, to showcase how it overcomes data privacy issues and enables knowledge sharing between alliance members in a Federated learning setting.

Via

Access Paper or Ask Questions

Scalable federated machine learning with FEDn

Feb 27, 2021

Morgan Ekmefjord, Addi Ait-Mlouk, Sadi Alawadi, Mattias Åkesson, Desislava Stoyanova, Ola Spjuth, Salman Toor, Andreas Hellander

Figure 1 for Scalable federated machine learning with FEDn

Figure 2 for Scalable federated machine learning with FEDn

Figure 3 for Scalable federated machine learning with FEDn

Figure 4 for Scalable federated machine learning with FEDn

Abstract:Federated machine learning has great promise to overcome the input privacy challenge in machine learning. The appearance of several projects capable of simulating federated learning has led to a corresponding rapid progress on algorithmic aspects of the problem. However, there is still a lack of federated machine learning frameworks that focus on fundamental aspects such as scalability, robustness, security, and performance in a geographically distributed setting. To bridge this gap we have designed and developed the FEDn framework. A main feature of FEDn is to support both cross-device and cross-silo training settings. This makes FEDn a powerful tool for researching a wide range of machine learning applications in a realistic setting.

Via

Access Paper or Ask Questions

Robust and integrative Bayesian neural networks for likelihood-free parameter inference

Feb 12, 2021

Fredrik Wrede, Robin Eriksson, Richard Jiang, Linda Petzold, Stefan Engblom, Andreas Hellander, Prashant Singh

Figure 1 for Robust and integrative Bayesian neural networks for likelihood-free parameter inference

Figure 2 for Robust and integrative Bayesian neural networks for likelihood-free parameter inference

Figure 3 for Robust and integrative Bayesian neural networks for likelihood-free parameter inference

Figure 4 for Robust and integrative Bayesian neural networks for likelihood-free parameter inference

Abstract:State-of-the-art neural network-based methods for learning summary statistics have delivered promising results for simulation-based likelihood-free parameter inference. Existing approaches require density estimation as a post-processing step building upon deterministic neural networks, and do not take network prediction uncertainty into account. This work proposes a robust integrated approach that learns summary statistics using Bayesian neural networks, and directly estimates the posterior density using categorical distributions. An adaptive sampling scheme selects simulation locations to efficiently and iteratively refine the predictive posterior of the network conditioned on observations. This allows for more efficient and robust convergence on comparatively large prior spaces. We demonstrate our approach on benchmark examples and compare against related methods.

Via

Access Paper or Ask Questions

Convolutional Neural Networks as Summary Statistics for Approximate Bayesian Computation

Jan 31, 2020

Mattias Åkesson, Prashant Singh, Fredrik Wrede, Andreas Hellander

Figure 1 for Convolutional Neural Networks as Summary Statistics for Approximate Bayesian Computation

Figure 2 for Convolutional Neural Networks as Summary Statistics for Approximate Bayesian Computation

Figure 3 for Convolutional Neural Networks as Summary Statistics for Approximate Bayesian Computation

Figure 4 for Convolutional Neural Networks as Summary Statistics for Approximate Bayesian Computation

Abstract:Approximate Bayesian Computation is widely used in systems biology for inferring parameters in stochastic gene regulatory network models. Its performance hinges critically on the ability to summarize high-dimensional system responses such as time series into a few informative, low-dimensional summary statistics. The quality of those statistics critically affect the accuracy of the inference. Existing methods to select the best subset out of a pool of candidate statistics do not scale well with large pools. Since it is imperative for good performance this becomes a serious bottleneck when doing inference on complex and high-dimensional problems. This paper proposes a convolutional neural network architecture for automatically learning informative summary statistics of temporal responses. We show that the proposed network can effectively circumvent the statistics selection problem as a preprocessing step to ABC for a challenging inference problem learning parameters in a high-dimensional stochastic genetic oscillator. We also study the impact of experimental design on network performance by comparing different data richness and different data acquisition strategies.

Via

Access Paper or Ask Questions