arXiv · 2312.15478
A Group Fairness Lens for Large Language Models
Abstract
The need to assess LLMs for bias and fairness is critical, with current evaluations often being narrow, missing a broad categorical view. In this paper, we propose evaluating the bias and fairness of LLMs from a group fairness lens using a novel hierarchical schema characterizing diverse social groups. Specifically, we construct a dataset, GFAIR, encapsulating target-attribute combinations across multiple dimensions. Moreover, we introduce statement organization, a new open-ended text generation task, to uncover complex biases in LLMs. Extensive evaluations of popular LLMs reveal inherent safety concerns. To mitigate the biases of LLMs from a group fairness perspective, we pioneer a novel chainof-thought method GF-THINK to mitigate biases of LLMs from a group fairness perspective. Experimental results demonstrate its efficacy in mitigating bias and achieving fairness in LLMs. Our dataset and codes are available at https://github.com/surika/Group-Fairness-LLMs.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Guanqun Bi, Yuqiang Xie, Lei Shen, Yanan Cao. 2023-12-24. A Group Fairness Lens for Large Language Models. https://arxiv.org/abs/2312.15478
Cite the original work for its findings. Save a collection to share your selection of sources.