Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Rafayel Darbinyan

In-context Learning in Presence of Spurious Correlations

Oct 04, 2024

Hrayr Harutyunyan, Rafayel Darbinyan, Samvel Karapetyan, Hrant Khachatrian

Figure 1 for In-context Learning in Presence of Spurious Correlations

Figure 2 for In-context Learning in Presence of Spurious Correlations

Figure 3 for In-context Learning in Presence of Spurious Correlations

Figure 4 for In-context Learning in Presence of Spurious Correlations

Abstract:Large language models exhibit a remarkable capacity for in-context learning, where they learn to solve tasks given a few examples. Recent work has shown that transformers can be trained to perform simple regression tasks in-context. This work explores the possibility of training an in-context learner for classification tasks involving spurious features. We find that the conventional approach of training in-context learners is susceptible to spurious features. Moreover, when the meta-training dataset includes instances of only one task, the conventional approach leads to task memorization and fails to produce a model that leverages context for predictions. Based on these observations, we propose a novel technique to train such a learner for a given classification task. Remarkably, this in-context learner matches and sometimes outperforms strong methods like ERM and GroupDRO. However, unlike these algorithms, it does not generalize well to other tasks. We show that it is possible to obtain an in-context learner that generalizes to unseen tasks by training on a diverse dataset of synthetic in-context learning instances.

Via

Access Paper or Ask Questions

Identifying and Disentangling Spurious Features in Pretrained Image Representations

Jun 22, 2023

Rafayel Darbinyan, Hrayr Harutyunyan, Aram H. Markosyan, Hrant Khachatrian

Figure 1 for Identifying and Disentangling Spurious Features in Pretrained Image Representations

Figure 2 for Identifying and Disentangling Spurious Features in Pretrained Image Representations

Figure 3 for Identifying and Disentangling Spurious Features in Pretrained Image Representations

Figure 4 for Identifying and Disentangling Spurious Features in Pretrained Image Representations

Abstract:Neural networks employ spurious correlations in their predictions, resulting in decreased performance when these correlations do not hold. Recent works suggest fixing pretrained representations and training a classification head that does not use spurious features. We investigate how spurious features are represented in pretrained representations and explore strategies for removing information about spurious features. Considering the Waterbirds dataset and a few pretrained representations, we find that even with full knowledge of spurious features, their removal is not straightforward due to entangled representation. To address this, we propose a linear autoencoder training method to separate the representation into core, spurious, and other features. We propose two effective spurious feature removal approaches that are applied to the encoding and significantly improve classification performance measured by worst group accuracy.

Via

Access Paper or Ask Questions

ML-based Approaches for Wireless NLOS Localization: Input Representations and Uncertainty Estimation

Apr 22, 2023

Rafayel Darbinyan, Hrant Khachatrian, Rafayel Mkrtchyan, Theofanis P. Raptis

Figure 1 for ML-based Approaches for Wireless NLOS Localization: Input Representations and Uncertainty Estimation

Figure 2 for ML-based Approaches for Wireless NLOS Localization: Input Representations and Uncertainty Estimation

Figure 3 for ML-based Approaches for Wireless NLOS Localization: Input Representations and Uncertainty Estimation

Figure 4 for ML-based Approaches for Wireless NLOS Localization: Input Representations and Uncertainty Estimation

Abstract:The challenging problem of non-line-of-sight (NLOS) localization is critical for many wireless networking applications. The lack of available datasets has made NLOS localization difficult to tackle with ML-driven methods, but recent developments in synthetic dataset generation have provided new opportunities for research. This paper explores three different input representations: (i) single wireless radio path features, (ii) wireless radio link features (multi-path), and (iii) image-based representations. Inspired by the two latter new representations, we design two convolutional neural networks (CNNs) and we demonstrate that, although not significantly improving the NLOS localization performance, they are able to support richer prediction outputs, thus allowing deeper analysis of the predictions. In particular, the richer outputs enable reliable identification of non-trustworthy predictions and support the prediction of the top-K candidate locations for a given instance. We also measure how the availability of various features (such as angles of signal departure and arrival) affects the model's performance, providing insights about the types of data that should be collected for enhanced NLOS localization. Our insights motivate future work on building more efficient neural architectures and input representations for improved NLOS localization performance, along with additional useful application features.

* This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible. Work partly supported by the RA Science Committee grant No. 22rl-052 (DISTAL) and the EU under Italian National Recovery and Resilience Plan of NextGenerationEU on "Telecommunications of the Future" (PE00000001 - program "RESTART")

Via

Access Paper or Ask Questions