Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Cornelia Sindermann

Prompt-based Personality Profiling: Reinforcement Learning for Relevance Filtering

Sep 06, 2024

Jan Hofmann, Cornelia Sindermann, Roman Klinger

Figure 1 for Prompt-based Personality Profiling: Reinforcement Learning for Relevance Filtering

Figure 2 for Prompt-based Personality Profiling: Reinforcement Learning for Relevance Filtering

Figure 3 for Prompt-based Personality Profiling: Reinforcement Learning for Relevance Filtering

Figure 4 for Prompt-based Personality Profiling: Reinforcement Learning for Relevance Filtering

Abstract:Author profiling is the task of inferring characteristics about individuals by analyzing content they share. Supervised machine learning still dominates automatic systems that perform this task, despite the popularity of prompting large language models to address natural language understanding tasks. One reason is that the classification instances consist of large amounts of posts, potentially a whole user profile, which may exceed the input length of Transformers. Even if a model can use a large context window, the entirety of posts makes the application of API-accessed black box systems costly and slow, next to issues which come with such "needle-in-the-haystack" tasks. To mitigate this limitation, we propose a new method for author profiling which aims at distinguishing relevant from irrelevant content first, followed by the actual user profiling only with relevant data. To circumvent the need for relevance-annotated data, we optimize this relevance filter via reinforcement learning with a reward function that utilizes the zero-shot capabilities of large language models. We evaluate our method for Big Five personality trait prediction on two Twitter corpora. On publicly available real-world data with a skewed label distribution, our method shows similar efficacy to using all posts in a user profile, but with a substantially shorter context. An evaluation on a version of these data balanced with artificial posts shows that the filtering to relevant posts leads to a significantly improved accuracy of the predictions.

* preprint, under review, supplementary material will be made available upon acceptance of the paper

Via

Access Paper or Ask Questions