arXiv · 1803.02077
The Contextual Loss for Image Transformation with Non-Aligned Data
Abstract
Feed-forward CNNs trained for image transformation problems rely on loss functions that measure the similarity between the generated image and a target image. Most of the common loss functions assume that these images are spatially aligned and compare pixels at corresponding locations. However, for many tasks, aligned training pairs of images will not be available. We present an alternative loss function that does not require alignment, thus providing an effective and simple solution for a new space of problems. Our loss is based on both context and semantics -- it compares regions with similar semantic meaning, while considering the context of the entire image. Hence, for example, when transferring the style of one face to another, it will translate eyes-to-eyes and mouth-to-mouth. Our code can be found at https://www.github.com/roimehrez/contextualLoss
Explore related subjects
Keep this discovery
Roey Mechrez, Itamar Talmi, Lihi Zelnik-Manor. 2018-03-06. The Contextual Loss for Image Transformation with Non-Aligned Data. https://arxiv.org/abs/1803.02077
Cite the original work for its findings. Save a collection to share your selection of sources.