Back to tools

Imagen

Imagen is a text-to-image diffusion model that achieves unprecedented photorealism and a deep level of language understanding. It uses large pretrained frozen text encoders and a new thresholding diffusion sampler to generate high-quality images.

ActiveFirst listed: Oct 3, 2024
Visit website

Best for

Users seeking text to image tools for their daily workflows.

Things to know

Usage limits, quotas, and premium capabilities are determined by the provider.

About this tool

Overview

Discover the power of Imagen, a text-to-image diffusion model that achieves unprecedented photorealism and a deep level of language understanding.

Large Pretrained Frozen Text Encoders

Imagen uses large pretrained frozen text encoders to generate high-quality images.

Thresholding Diffusion Sampler

Imagen uses a new thresholding diffusion sampler to generate high-quality images.

Efficient U-Net Architecture

Imagen uses an efficient U-Net architecture that is more compute efficient, more memory efficient, and converges faster.

Cascaded Diffusion Models

Imagen uses cascaded diffusion models to generate high-resolution images.

Get started

  1. Open the official website and confirm the service is available in your region.
  2. Check the current plan, usage limits and terms for your intended use.
  3. Try a small task with sample data before committing to a paid plan.

Editorial note

Pricing checked: Sep 20, 2026. Official website is active. Please verify current pricing and terms directly on the official site.

Record updated: Sep 20, 2026

Suggest a correction