Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Aneesha Sampath

Beyond Binary: Multiclass Paraphasia Detection with Generative Pretrained Transformers and End-to-End Models

Jul 16, 2024

Matthew Perez, Aneesha Sampath, Minxue Niu, Emily Mower Provost

Abstract:Aphasia is a language disorder that can lead to speech errors known as paraphasias, which involve the misuse, substitution, or invention of words. Automatic paraphasia detection can help those with Aphasia by facilitating clinical assessment and treatment planning options. However, most automatic paraphasia detection works have focused solely on binary detection, which involves recognizing only the presence or absence of a paraphasia. Multiclass paraphasia detection represents an unexplored area of research that focuses on identifying multiple types of paraphasias and where they occur in a given speech segment. We present novel approaches that use a generative pretrained transformer (GPT) to identify paraphasias from transcripts as well as two end-to-end approaches that focus on modeling both automatic speech recognition (ASR) and paraphasia classification as multiple sequences vs. a single sequence. We demonstrate that a single sequence model outperforms GPT baselines for multiclass paraphasia detection.

Via

Access Paper or Ask Questions

SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Nov 23, 2022

Aneesha Sampath, Victoria Lin, Louis-Philippe Morency

Figure 1 for SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Figure 2 for SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Figure 3 for SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Figure 4 for SeedBERT: Recovering Annotator Rating Distributions from an Aggregated Label

Abstract:Many machine learning tasks -- particularly those in affective computing -- are inherently subjective. When asked to classify facial expressions or to rate an individual's attractiveness, humans may disagree with one another, and no single answer may be objectively correct. However, machine learning datasets commonly have just one "ground truth" label for each sample, so models trained on these labels may not perform well on tasks that are subjective in nature. Though allowing models to learn from the individual annotators' ratings may help, most datasets do not provide annotator-specific labels for each sample. To address this issue, we propose SeedBERT, a method for recovering annotator rating distributions from a single label by inducing pre-trained models to attend to different portions of the input. Our human evaluations indicate that SeedBERT's attention mechanism is consistent with human sources of annotator disagreement. Moreover, in our empirical evaluations using large language models, SeedBERT demonstrates substantial gains in performance on downstream subjective tasks compared both to standard deep learning models and to other current models that account explicitly for annotator disagreement.

* To be published in AAAI-23 Workshop on Uncertainty Reasoning and Quantification in Decision Making

Via

Access Paper or Ask Questions