Papers › iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow...

iGAiVA: Integrated Generative AI and Visual Analytics in a Machine Learning Workflow for Text Classification

24 Sep 2024arXiv:2409.15848archive 2025-07-28

Yuanzhe Jin, Adrian Carrasco-Revilla, Min Chen

In developing machine learning (ML) models for text classification, one common challenge is that the collected data is often not ideally distributed, especially when new classes are introduced in response to changes of data and tasks. In this paper, we present a solution for using visual analytics (VA) to guide the generation of synthetic data using large language models. As VA enables model developers to identify data-related deficiency, data synthesis can be targeted to address such deficiency. We discuss different types of data deficiency, describe different VA techniques for supporting their identification, and demonstrate the effectiveness of targeted data synthesis in improving model accuracy. In addition, we present a software tool, iGAiVA, which maps four groups of ML tasks into four VA views, integrating generative AI and VA into an ML workflow for developing and improving text classification models.

PaperPDFCode

Code

mattjin19/rbf officialmentioned in papermentioned on GitHub report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Text Classificationtext-classification

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

Visual Analytics

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections