TY - RPRT TI - Cycle Consistency as Reward: Learning Image-Text Alignment without Human Preferences AU - Hyojin Bahng AU - Caroline Chan AU - Fredo Durand AU - Phillip Isola PY - 2025 UR - https://arxiv.org/abs/2506.02095 ID - 2506.02095 ER -