Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding project page: sota FID(7.27 on COCO), without ever training on COCO, human raters find Imagen samples to be on par with the COCO data itself in image-text alignmenthttps://foxdeo.com/TYme.tnk