arXiv · 2409.03307
AI data transparency: an exploration through the lens of AI incidents
Abstract
Knowing more about the data used to build AI systems is critical for allowing different stakeholders to play their part in ensuring responsible and appropriate deployment and use. Meanwhile, a 2023 report shows that data transparency lags significantly behind other areas of AI transparency in popular foundation models. In this research, we sought to build on these findings, exploring the status of public documentation about data practices within AI systems generating public concern. Our findings demonstrate that low data transparency persists across a wide range of systems, and further that issues of transparency and explainability at model- and system- level create barriers for investigating data transparency information to address public concerns about AI systems. We highlight a need to develop systematic ways of monitoring AI data transparency that account for the diversity of AI system types, and for such efforts to build on further understanding of the needs of those both supplying and using data transparency information.
Explore related subjects
Keep this discovery
Sophia Worth, Ben Snaith, Arunav Das, Gefion Thuermer, Elena Simperl. 2024-09-05. AI data transparency: an exploration through the lens of AI incidents. https://arxiv.org/abs/2409.03307
Cite the original work for its findings. Save a collection to share your selection of sources.