arXiv · 2207.13744
Lighting (In)consistency of Paint by Text
Abstract
Whereas generative adversarial networks are capable of synthesizing highly realistic images of faces, cats, landscapes, or almost any other single category, paint-by-text synthesis engines can -- from a single text prompt -- synthesize realistic images of seemingly endless categories with arbitrary configurations and combinations. This powerful technology poses new challenges to the photo-forensic community. Motivated by the fact that paint by text is not based on explicit geometric or physical models, and the human visual system's general insensitivity to lighting inconsistencies, we provide an initial exploration of the lighting consistency of DALL-E-2 synthesized images to determine if physics-based forensic analyses will prove fruitful in detecting this new breed of synthetic media.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hany Farid. 2022-07-27. Lighting (In)consistency of Paint by Text. https://arxiv.org/abs/2207.13744
Cite the original work for its findings. Save a collection to share your selection of sources.