arXiv · 2603.09979
GhazalBench: Canonical Verse Access in LLMs across Persian Ghazals and Shakespearean Sonnets
Abstract
Persian poetry plays an active role in Iranian cultural practice, where verses by canonical poets such as Hafez and Saadi are frequently quoted, paraphrased, or completed from incomplete cues. Supporting such interactions requires language models to reliably access canonical verses from semantic and lexical information. We introduce GhazalBench, a benchmark for evaluating how large language models (LLMs) access the canonical surface forms of Persian ghazal under usage-grounded conditions. Unlike prior work that primarily treats memorization as a liability, GhazalBench studies settings in which access to exact wording is functionally important. The benchmark evaluates completion and recognition under varied semantic and lexical cues. Across proprietary and open-weight multilingual LLMs, exact completion remains challenging, while recognition is substantially stronger. Performance also varies across poets and model families. Parallel experiments on Shakespearean sonnets yield markedly higher completion for several models, consistent with differences in exposure to canonical texts. {Additional experiments suggest that post-training may reduce direct continuation-based access to canonical text.} Our findings motivate evaluation frameworks that distinguish production from recognition and assess access to culturally significant canonical texts. GhazalBench is available at https://github.com/kalhorghazal/GhazalBench.
Explore related subjects
Keep this discovery
Ghazal Kalhor, Yadollah Yaghoobzadeh. 2026-02-06. GhazalBench: Canonical Verse Access in LLMs across Persian Ghazals and Shakespearean Sonnets. https://arxiv.org/abs/2603.09979
Cite the original work for its findings. Save a collection to share your selection of sources.