TY - RPRT TI - Benchmarking Adversarial Robustness to Bias Elicitation in Large Language Models: Scalable Automated Assessment with LLM-as-a-Judge AU - Riccardo Cantini AU - Alessio Orsino AU - Massimo Ruggiero AU - Domenico Talia PY - 2025 DO - 10.1007/s10994-025-06862-6 UR - https://arxiv.org/abs/2504.07887 ID - 2504.07887 ER -