Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Adrian Carrasco-Revilla

iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

Sep 24, 2024

Yuanzhe Jin, Adrian Carrasco-Revilla, Min Chen

Figure 1 for iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

Figure 2 for iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

Figure 3 for iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

Figure 4 for iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

Abstract:In developing machine learning (ML) models for text classification, one common challenge is that the collected data is often not ideally distributed, especially when new classes are introduced in response to changes of data and tasks. In this paper, we present a solution for using visual analytics (VA) to guide the generation of synthetic data using large language models. As VA enables model developers to identify data-related deficiency, data synthesis can be targeted to address such deficiency. We discuss different types of data deficiency, describe different VA techniques for supporting their identification, and demonstrate the effectiveness of targeted data synthesis in improving model accuracy. In addition, we present a software tool, iGAiVA, which maps four groups of ML tasks into four VA views, integrating generative AI and VA into an ML workflow for developing and improving text classification models.

Via

Access Paper or Ask Questions