Google Research has introduced Imagen, a text-to-image diffusion model that achieves unprecedented photorealism and language understanding. Imagen uses a large frozen language model to encode input text and a cascaded diffusion model to generate high-fidelity images. The model has achieved a new state-of-the-art FID score on the COCO dataset and has been evaluated through a new benchmark called DrawBench. However, the researchers have decided not to release the model due to its potential social biases and limitations.