Score: 0

Small sample-based adaptive text classification through iterative and contrastive description refinement

Published: August 1, 2025 | arXiv ID: 2508.00957v1

By: Amrit Rajeev , Udayaadithya Avadhanam , Harshula Tulapurkar and more

Potential Business Impact:

Teaches computers to sort text without new training.

Zero-shot text classification remains a difficult task in domains with evolving knowledge and ambiguous category boundaries, such as ticketing systems. Large language models (LLMs) often struggle to generalize in these scenarios due to limited topic separability, while few-shot methods are constrained by insufficient data diversity. We propose a classification framework that combines iterative topic refinement, contrastive prompting, and active learning. Starting with a small set of labeled samples, the model generates initial topic labels. Misclassified or ambiguous samples are then used in an iterative contrastive prompting process to refine category distinctions by explicitly teaching the model to differentiate between closely related classes. The framework features a human-in-the-loop component, allowing users to introduce or revise category definitions in natural language. This enables seamless integration of new, unseen categories without retraining, making the system well-suited for real-world, dynamic environments. The evaluations on AGNews and DBpedia demonstrate strong performance: 91% accuracy on AGNews (3 seen, 1 unseen class) and 84% on DBpedia (8 seen, 1 unseen), with minimal accuracy shift after introducing unseen classes (82% and 87%, respectively). The results highlight the effectiveness of prompt-based semantic reasoning for fine-grained classification with limited supervision.

LLM-as-classifier: Semi-Supervised, Iterative Framework for Hierarchical Text Classification using Large Language Models

Computation and Language

Makes smart computer programs sort text better.

22 Aug 2025 0

89%

Bridging the Gap: In-Context Learning for Modeling Human Disagreement

Computation and Language

Helps computers understand when people disagree.

6 Jun 2025 1

88%

Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study

Computers and Society

Helps surveys fill in missing answers accurately.

29 Sep 2025 1

View PDF Login to Bookmark

Page Count

12 pages

Small sample-based adaptive text classification through iterative and contrastive description refinement

Teaches computers to sort text without new training.

Technical Abstract

LLM-as-classifier: Semi-Supervised, Iterative Framework for Hierarchical Text Classification using Large Language Models

Bridging the Gap: In-Context Learning for Modeling Human Disagreement

Learning from Convenience Samples: A Case Study on Fine-Tuning LLMs for Survey Non-response in the German Longitudinal Election Study