arXiv · 2511.20667
A centroid based framework for text classification in itsm environments
Abstract
Text classification with hierarchical taxonomies is a fundamental requirement in IT Service Management (ITSM) systems, where support tickets must be categorized into tree-structured taxonomies. We present a dual-embedding centroid-based classification framework that maintains separate semantic and lexical centroid representations per category, combining them through reciprocal rank fusion at inference time. The framework achieves performance competitive with Support Vector Machines (hierarchical F1: 0.731 vs 0.727) while providing interpretability through centroid representations. Evaluated on 8,968 ITSM tickets across 123 categories, this method achieves 5.9 times faster training and up to 152 times faster incremental updates. With 8.6-8.8 times speedup across batch sizes (100-1000 samples) when excluding embedding computation. These results make the method suitable for production ITSM environments prioritizing interpretability and operational efficiency.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hossein Mohanna, Ali Ait-Bachir. 2025-11-12. A centroid based framework for text classification in itsm environments. https://arxiv.org/abs/2511.20667
Cite the original work for its findings. Save a collection to share your selection of sources.