arXiv · 2510.15685
Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech Detection
Abstract
This paper investigates the use of an LLM to generate auxiliary background context for social media posts, and explores four methods to incorporate this context into the input of an SBERT-based Hate Speech Detection (HSD) classifier. These are: text concatenation, embedding concatenation, a hierarchical transformer-based fusion, and LLM-driven text enhancement. We evaluate the impact of our context generation and incorporation strategies in a textual setting on the Latent Hatred dataset of implicitly hateful tweets and a multimodal setting on the MAMI dataset of misogynous internet memes. Results are evaluated against a zero-context baseline, two previous approaches based on entity linking, and a zero-shot LLM classifier. Findings indicate that incorporating generated context improves HSD performance by up to 3 and 6 F1 points on textual and multimodal settings respectively, from a zero-context baseline to the highest-performing system, based on embedding concatenation.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Joshua Wolfe Brook, Ilia Markov. 2025-10-17. Leveraging LLMs for Context-Aware Implicit Textual and Multimodal Hate Speech Detection. https://arxiv.org/abs/2510.15685
Cite the original work for its findings. Save a collection to share your selection of sources.