arXiv · 2505.18215
Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?
Abstract
The rapid adoption of LLMs has overshadowed the potential advantages of traditional BERT-like models in text classification. This study challenges the prevailing "LLM-centric" trend by systematically comparing three category methods, i.e., BERT-like models fine-tuning, LLM internal state utilization, and zero-shot inference across six high-difficulty datasets. Our findings reveal that BERT-like models often outperform LLMs. We further categorize datasets into three types, perform PCA and probing experiments, and identify task-specific model strengths: BERT-like models excel in pattern-driven tasks, while LLMs dominate those requiring deep semantics or world knowledge. Based on this, we propose TaMAS, a fine-grained task selection strategy, advocating for a nuanced, task-driven approach over a one-size-fits-all reliance on LLMs.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Junyan Zhang, Yiming Huang, Shuliang Liu, Yubo Gao, Xuming Hu. 2025-05-23. Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?. https://arxiv.org/abs/2505.18215
Cite the original work for its findings. Save a collection to share your selection of sources.