arXiv · 2507.10468
From BERT to Qwen: Hate Detection across architectures
Abstract
Online platforms struggle to curb hate speech without over-censoring legitimate discourse. Early bidirectional transformer encoders made big strides, but the arrival of ultra-large autoregressive LLMs promises deeper context-awareness. Whether this extra scale actually improves practical hate-speech detection on real-world text remains unverified. Our study puts this question to the test by benchmarking both model families, classic encoders and next-generation LLMs, on curated corpora of online interactions for hate-speech detection (Hate or No Hate).
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ariadna Mon, Saúl Fenollosa, Jon Lecumberri. 2025-07-14. From BERT to Qwen: Hate Detection across architectures. https://arxiv.org/abs/2507.10468
Cite the original work for its findings. Save a collection to share your selection of sources.