Clever little tricks: A socio-technical history of text-to-image generative models

The emergence of text-to-image generative models (e.g., Midjourney, DALL-E 2, Stable Diffusion) in the summer of 2022 impacted architectural visual culture suddenly, severely, and seemingly out of nowhere. To contextualize this phenomenon, this text offers a socio-technical history of text-to-image...

Full description

Saved in:

Bibliographic Details
Published in	International journal of architectural computing Vol. 21; no. 2; pp. 211 - 241
Main Author	Steinfeld, Kyle
Format	Journal Article
Language	English
Published	London, England SAGE Publications 01.06.2023
Subjects	socio-technical study text-to-image generative AI Machine learning
Online Access	Get full text

Cover

Loading…

More Information
Summary:	The emergence of text-to-image generative models (e.g., Midjourney, DALL-E 2, Stable Diffusion) in the summer of 2022 impacted architectural visual culture suddenly, severely, and seemingly out of nowhere. To contextualize this phenomenon, this text offers a socio-technical history of text-to-image generative systems. Three moments in time, or “scenes,” are presented here: the first at the advent of AI in the middle of the last century; the second at the “reawakening” of a specific approach to machine learning at the turn of this century; the third that documents a rapid sequence of innovations, dubbed “clever little tricks,” that occurred across just 18 months. This final scene is the crux, and represents the first formal documentation of the recent history of a specific set of informal innovations. These innovations were produced by non-affiliated researchers and communities of creative contributors, and directly led to the technologies that so compellingly captured the architectural imagination in the summer of 2022. Across these scenes, we examine the technologies, application domains, infrastructures, social contexts, and practices that drive technical research and shape creative practice in this space.
ISSN:	1478-0771 2048-3988
DOI:	10.1177/14780771231168230