You could train text-to-image models without expensive human feedback! Alibaba Group and Zhejiang University researchers present PromptEcho—a reward method that uses a frozen vision-language model to measure image-prompt alignment directly, with zero annotations or extra
PromptEcho: A New Zero-Annotation Reward Method for Text-to-Image Models
By
–
