ANTI-AI ARCHIVE
ART-HISTORY NODE / 083 · 2022-05

Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Chitwan Saharia et al. / Google Research

Historical context & research account

Imagen uses a large text encoder and diffusion decoders. The research highlights gains in image–text alignment from scaling the language model. It develops a direction that later multimodal systems also pursue: improvements in understanding language can change the visual capabilities and controllability of image generation.

Dates & version record

Date displayed for this node:2022-05 · Date recorded in the research master:2022-05

These dates refer to the historical event or recorded version, not this page’s publication date. The original date precision and unresolved questions are retained.

Original sources & further reading

These links lead to the cited paper, article, institution or conference page. External texts retain their source languages.

01
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding ↗

Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

https://arxiv.org/abs/2205.11487

Current cited URL

02
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding ↗

Imagen

https://imagen.research.google/

URL cited in a research master or supplement; it may have moved or expired

Provenance, translation & verification

Archive node #083 · Historical node

Research masters & supplement references · 3

AI Art Genealogy Research Master · p. 25 · 083
Date as recorded: 2022-05
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding ↗
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

GenAI Visual Release Chronology Research Master · p. 5 · R020
Date as recorded: 2022-05
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding ↗
Google Imagen

GenAI Visual Release Chronology Research Master · p. 6 · R021
Date as recorded: 2022-05
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding ↗
Imagen

Archive account adapted from the masters, not a full translation of the linked work. AI-assisted translation; human review pending.

The master’s research claims and dates await independent verification.

Cite this node

ANTI-AI ARCHIVE. Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding. Art-history node #083. https://salondesrefuses.cn/en/art-history/083

For specific historical claims, also cite the original sources above and include your access date. This account is not a full translation of the linked work.

Adjacent nodes follow chronological order; adjacency does not establish direct influence or causation.